arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

自动驾驶

自动驾驶感知、规划、BEV、占用预测、激光雷达和仿真评测。

2025-12-24 至 2025-12-24 共收录 13 信号源:cs.RO, cs.CV, eess.IV, cs.AI

1. 感知 4 篇

2507.20538 2025-12-24 cs.RO 70%

Uni-Mapper: Unified Mapping Framework for Multi-modal LiDARs in Complex and Dynamic Environments

Uni-Mapper:多模态激光雷达在复杂动态环境中的统一映射框架

Gilhwan Kang, Hogyun Kim, Byunghee Choi, Seokhwan Jeong, Young-Sik Shin, Younggun Cho

机构 * Hyundai Motor Company(现代汽车公司) Inha University(inha大学) Korea Institute of Machinery and Materials(韩国机械材料研究院)

专题命中 感知 :LiDAR(abstract);occupancy(abstract);分类 cs.RO

AI总结 Uni-Mapper通过动态感知和多模态激光雷达融合技术,实现复杂动态环境下的统一地图构建与回环检测。

Comments 18 pages, 14 figures

Journal ref 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.19934 2025-12-24 cs.CV cs.AI cs.LG 62%

Vehicle-centric Perception via Multimodal Structured Pre-training

基于多模态结构预训练的车辆感知

Wentao Wu, Xiao Wang, Chenglong Li, Jin Tang, Bin Luo

机构 * Information Materials and Intelligent Sensing Laboratory of Anhui Province(安徽省信息材料与智能感知实验室) Anhui Provincial Key Laboratory of Multimodal Cognitive Computation(安徽省多模态认知计算重点实验室) the School of Artificial Intelligence, Anhui University(安徽大学人工智能学院) School of Computer Science and Technology, Anhui University(安徽大学计算机科学与技术学院) Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合性国家科学中心人工智能研究院)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

AI总结 本文提出VehicleMAE-V2,通过多模态结构先验知识提升车辆感知的预训练能力,采用SMM、CRM和SRM模块增强模型对车辆结构和语义的理解。

Comments Journal extension of VehicleMAE (AAAI 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.20217 2025-12-24 cs.CV 57%

LiteFusion: Taming 3D Object Detectors from Vision-Based to Multi-Modal with Minimal Adaptation

LiteFusion: 从基于视觉到多模态的3D目标检测器的最小适应

Xiangxuan Ren, Zhongdao Wang, Pin Tang, Guoqing Wang, Jilai Zheng, Chao Ma

机构 * China Ministry of Education (MOE) Key Laboratory of Artificial Intelligence(中国教育部人工智能重点实验室) Artificial Intelligence Institute, Shanghai Jiao Tong University(上海交通大学人工智能学院) Huawei Noah’s Ark Lab(华为诺亚实验室)

专题命中 感知 :LiDAR(abstract);分类 cs.CV

AI总结 LiteFusion通过将LiDAR数据作为补充几何信息源,无需专用LiDAR编码器,显著提升多模态3D目标检测性能。

Comments 13 pages, 9 figures, 8 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.19725 2025-12-24 cs.LG 50%

Out-of-Distribution Detection for Continual Learning: Design Principles and Benchmarking

持续学习中的分布外检测:设计原则与基准测试

Srishti Gupta, Riccardo Balia, Daniele Angioni, Fabio Brau, Maura Pintor, Ambra Demontis, Alessandro Sebastian, Salvatore Mario Carta, Fabio Roli, Battista Biggio

专题命中 感知 :autonomous driving(abstract)

AI总结 本文探讨了持续学习中分布外检测的设计原则与基准测试,旨在提升AI系统在动态环境中的适应性和鲁棒性。

Comments International Journal of Computer Vision

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 规划控制 3 篇

2507.05710 2025-12-24 cs.RO 79%

DRO-EDL-MPC: Evidential Deep Learning-Based Distributionally Robust Model Predictive Control for Safe Autonomous Driving

DRO-EDL-MPC:基于证据深度学习的分布鲁棒模型预测控制用于安全自动驾驶

Hyeongchan Ham, Heejin Ahn

机构 * School of Electrical Engineering, Korea Advanced Institute of Science and Technology (KAIST)(电气工程学院,韩国科学技术院)

专题命中 规划控制 :autonomous driving(title,abstract);分类 cs.RO

AI总结 DRO-EDL-MPC通过结合证据深度学习和分布鲁棒优化,提升自动驾驶中的安全性和控制可靠性。

Comments Accepted to IEEE Robotics and Automation Letters (RA-L)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.19629 2025-12-24 cs.RO cs.CV 62%

LoGoPlanner: Localization Grounded Navigation Policy with Metric-aware Visual Geometry

LoGoPlanner:基于度量感知的视觉几何导航策略

Jiaqi Peng, Wenzhe Cai, Yuqiang Yang, Tai Wang, Yuan Shen, Jiangmiao Pang

机构 * Department of Electronic Engineering, Tsinghua University(清华大学电子工程系) Shanghai AI laboratory(上海人工智能实验室)

专题命中 规划控制 :trajectory planning(abstract);分类 cs.RO、cs.CV

AI总结 LoGoPlanner通过结合度量感知的视觉几何和端到端学习,提升移动机器人在无结构环境中的导航性能和泛化能力。

Comments Project page:https://steinate.github.io/logoplanner.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.20425 2025-12-24 physics.optics 50%

Arbitrary laser frequency modulation algorithm based on iterative on-the-fly deconvolution

基于迭代即席反卷积的任意激光频率调制算法

Thierry Chanelière

专题命中 规划控制 :LiDAR(abstract)

AI总结 本文提出了一种适用于任意调制模式的激光频率调制算法,通过迭代反卷积技术实现即席处理,并通过实验验证了其在FMCW和正方形波调制中的有效性。

Journal ref Appl. Opt. 65, 341-348 (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏

3. BEV与占用 3 篇

2507.13387 2025-12-24 cs.CV eess.IV 86%

From Binary to Semantic: Utilizing Large-Scale Binary Occupancy Data for 3D Semantic Occupancy Prediction

从二元到语义:利用大规模二元占用数据进行3D语义占用预测

Chihiro Noguchi, Takaki Yamamoto

机构 * InfoTech, Toyota Motor Corporation(丰田汽车公司信息科技部)

专题命中 BEV与占用 :occupancy(title,abstract);autonomous driving(abstract);LiDAR(abstract);分类 cs.CV、eess.IV

AI总结 本文提出一种基于二元占用数据的框架,通过预训练和自动标注方法提升3D语义占用预测的性能。

Comments Accepted to ICCV Workshop 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.19871 2025-12-24 cs.CV 83%

HyGE-Occ: Hybrid View-Transformation with 3D Gaussian and Edge Priors for 3D Panoptic Occupancy Prediction

HyGE-Occ: 基于3D高斯和边缘先验的混合视图变换用于3D全景占用预测

Jong Wook Kim, Wonseok Roh, Ha Dam Baek, Pilhyeon Lee, Jonghyun Choi, Sangpil Kim

机构 * Korea University(韩国大学) Hyundai Motor Company(现代汽车公司) Inha University(inha大学)

专题命中 BEV与占用 :occupancy(title,abstract);BEV(abstract);分类 cs.CV

AI总结 HyGE-Occ通过结合3D高斯和边缘先验的混合视图变换,提升3D全景占用预测的几何一致性和边界意识,实现更精确的3D场景重建。

Comments 11 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18279 2025-12-24 cs.CV 70%

UniMPR: A Unified Framework for Multimodal Place Recognition with Heterogeneous Sensor Configurations

UniMPR: 一种用于异构传感器配置的多模态地点识别统一框架

Zhangshuo Qi, Jingyi Xu, Luqi Cheng, Shichen Wen, Yiming Ma, Guangming Xiong

机构 * Beijing Institute of Technology(北京理工大学) Shanghai Jiao Tong University(上海交通大学) The University of New South Wales(新南威尔士大学)

专题命中 BEV与占用 :BEV(abstract);LiDAR(abstract);分类 cs.CV

AI总结 UniMPR提出了一种统一框架,能够适应多种异构传感器配置,通过极坐标BEV特征空间和多分支网络实现多模态地点识别的高效与鲁棒性。

Comments 14 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 激光雷达 1 篇

2512.20224 2025-12-24 cs.RO 70%

UrbanV2X: A Multisensory Vehicle-Infrastructure Dataset for Cooperative Navigation in Urban Areas

UrbanV2X: 一种用于城市区域协同导航的多感官车辆-基础设施数据集

Qijun Qin, Ziqi Zhang, Yihan Zhong, Feng Huang, Xikun Liu, Runzhi Hu, Hang Chen, Wei Hu, Dongzhe Su, Jun Zhang, Hoi-Fung Ng, Weisong Wen

机构 * Department of Aeronautical and Aviation Engineering, The Hong Kong Polytechnic University(航空与航空工程系,香港理工大学) Hong Kong Applied Science and Technology Research Institute (ASTRI)(香港应用科学和技术研究院) Nanyang Technological University(南洋理工大学)

专题命中 激光雷达 :autonomous driving(abstract);LiDAR(abstract);分类 cs.RO

AI总结 UrbanV2X数据集通过多感官数据支持城市区域的协同导航研究,包含车辆和基础设施的同步传感器数据及算法评估

Comments 8 pages, 9 figures, IEEE ITSC 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 仿真评测 1 篇

2512.17586 2025-12-24 cs.LG cs.RO 79%

Learning Safe Autonomous Driving Policies Using Predictive Safety Representations

利用预测安全表示学习安全的自动驾驶策略

Mahesh Keswani, Raunak Bhattacharyya

机构 * Yardi School of Artificial Intelligence, Indian Institute of Technology Delhi(雅迪人工智能学院,印度理工学院德里)

专题命中 仿真评测 :autonomous driving(title,abstract);分类 cs.RO

AI总结 本文提出利用预测安全表示改进自动驾驶策略学习,通过实验验证了其在提升安全性和效率方面的有效性。

Comments 8 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

6. 其他自动驾驶 1 篇

2512.20179 2025-12-24 cs.HC 50%

RESPOND: Risk-Enhanced Structured Pattern for LLM-driven Online Node-level Decision-making

基于风险增强的结构化模式:面向LLM驱动的在线节点级决策

Dan Chen, Heye Huang, Tiantian Chen, Zheng Li, Yongji Li, Yuhui Xu, Sikai Chen

专题命中 其他自动驾驶 :autonomous driving(abstract)

AI总结 RESPOND通过结构化模式和风险增强机制提升LLM驱动驾驶代理的决策精度与安全性,有效减少碰撞并实现个性化驾驶风格适应。

Comments 28 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏