arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

自动驾驶

自动驾驶感知、规划、BEV、占用预测、激光雷达和仿真评测。

2026-02-09 至 2026-02-09 共收录 13 信号源:cs.RO, cs.CV, eess.IV, cs.AI

1. 感知 5 篇

2106.10823 2026-02-09 cs.CV 79%

3D Object Detection for Autonomous Driving: A Survey

自动驾驶中的3D物体检测:综述

Rui Qian, Xin Lai, Xirong Li

机构 * Key Lab of Data Engineering and Knowledge Engineering, Renmin University of China, Beijing 100872, China(数据工程与知识工程重点实验室,中国人民大学,北京100872,中国) School of Mathematics, Renmin University of China, Beijing 100872, China(数学学院,中国人民大学,北京100872,中国)

专题命中 感知 :autonomous driving(title,abstract);分类 cs.CV

AI总结 本文综述了自动驾驶中3D物体检测的研究进展,涵盖了传感器、数据集、性能指标及最新方法,并分析了其优缺点及未来方向。

Comments The manuscript is accepted by Pattern Recognition on 14 May 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.06406 2026-02-09 cs.CV 70%

Point Virtual Transformer

点虚拟变换器

Veerain Sood, Bnalin, Gaurav Pandey

机构 * Texas A \& M University Email Engineering Technology \& Industrial Distribution Texas A \& M University Texas, USA

专题命中 感知 :BEV(abstract);LiDAR(abstract);分类 cs.CV

AI总结 PointViT通过融合真实和虚拟点,提升远距离3D目标检测性能,实现91.16%的3D AP和95.94%的BEV AP。

Comments 8 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.05538 2026-02-09 cs.CV 70%

A Comparative Study of 3D Person Detection: Sensor Modalities and Robustness in Diverse Indoor and Outdoor Environments

三维人物检测的比较研究:传感器模态与在多样室内和室外环境中的鲁棒性

Malaz Tamim, Andrea Matic-Flierl, Karsten Roscher

机构 * Fraunhofer Institute for Cognitive Systems IKS(弗劳恩霍夫认知系统研究所)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 本文比较了三种三维人物检测方法,发现融合模型在复杂场景中表现最佳,但对传感器错位仍敏感。

Comments Accepted for VISAPP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.04692 2026-02-09 cs.CV cs.AI 62%

DRMOT: A Dataset and Framework for RGBD Referring Multi-Object Tracking

DRMOT:用于RGBD参照多目标跟踪的数据集和框架

Sijia Chen, Lijuan Ma, Yanqiu Yu, En Yu, Liman Liu, Wenbing Tao

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

AI总结 本文提出DRMOT任务,通过融合RGB、深度和语言信息,构建DRSet数据集并提出DRTrack框架,提升多目标跟踪的3D感知能力。

Comments https://github.com/chen-si-jia/DRMOT

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.10330 2026-02-09 cs.CV 57%

BADet: Boundary-Aware 3D Object Detection from Point Clouds

BADet: 基于点云的边界感知3D目标检测

Rui Qian, Xin Lai, Xirong Li

机构 * Key Lab of Data Engineering and Knowledge Engineering(数据工程与知识工程重点实验室) Renmin University of China(中国人民大学) School of Mathematics, Renmin University of China(中国人民大学数学学院) Department of Physics, J.K. Institute of Science(科学研究院物理系)

专题命中 感知 :BEV(abstract);分类 cs.CV

AI总结 BADet通过构建局部邻域图和轻量级特征聚合模块,提升基于点云的3D目标检测性能。

Comments The manuscript is accepted by Pattern Recognition on 6 Jan, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 规划控制 2 篇

2404.09657 2026-02-09 cs.RO cs.LG 88%

Sampling for Model Predictive Trajectory Planning in Autonomous Driving using Normalizing Flows

基于归一化流的自动驾驶模型预测轨迹规划采样

Georg Rabenstein, Lars Ullrich, Knut Graichen

机构 * IEEE

专题命中 规划控制 :autonomous driving(title,abstract);trajectory planning(title,abstract);分类 cs.RO

AI总结 本文提出基于归一化流的采样方法,用于提高自动驾驶轨迹规划的效率和探索能力。

Comments Accepted to be published as part of the 2024 IEEE Intelligent Vehicles Symposium (IV), Jeju Shinhwa World, Jeju Island, Korea, June 2-5, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07210 2026-02-09 cs.RO cs.AI 81%

HyPlan: Hybrid Learning-Assisted Planning Under Uncertainty for Safe Autonomous Driving

HyPlan:在不确定环境下通过混合学习辅助规划实现安全自动驾驶

Donald Pfaffmann, Matthias Klusch, Marcel Steinmetz

机构 * Saarland University, Computer Science Dept.(萨尔兰大学计算机科学系) German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心) French National Centre for Scientific Research (CNRS)(法国国家科学研究中心)

专题命中 规划控制 :autonomous driving(title);self-driving(abstract);分类 cs.RO、cs.AI

AI总结 HyPlan通过结合多智能体行为预测、深度强化学习和POMDP规划,实现安全且高效的自动驾驶导航

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 端到端驾驶 1 篇

2602.06521 2026-02-09 cs.CV cs.RO 81%

DriveWorld-VLA: Unified Latent-Space World Modeling with Vision-Language-Action for Autonomous Driving

DriveWorld-VLA: 基于视觉-语言-动作的统一潜在空间世界建模用于自动驾驶

Feiyang jia, Lin Liu, Ziying Song, Caiyan Jia, Hangjun Ye, Xiaoshuai Hao, Long Chen

机构 * School of Computer Science(计算机科学学院) Beijing Key Laboratory of Traffic Data Mining(交通数据挖掘重点实验室) Beijing Jiaotong University(北京交通大学)

专题命中 端到端驾驶 :autonomous driving(title,abstract);分类 cs.RO、cs.CV

AI总结 DriveWorld-VLA通过统一视觉-语言-动作与潜在空间世界建模,提升自动驾驶中的决策与前瞻性想象能力。

Comments 20 pages, 7 tables, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

4. BEV与占用 3 篇

2602.06488 2026-02-09 cs.CV 83%

Rebenchmarking Unsupervised Monocular 3D Occupancy Prediction

重新基准测试无监督单目3D占用预测

Zizhan Guo, Yi Feng, Mengtan Zhang, Haoran Zhang, Wei Ye, Rui Fan

专题命中 BEV与占用 :occupancy(title,abstract);autonomous driving(abstract);分类 cs.CV

AI总结 本文提出重新基准测试无监督单目3D占用预测,通过改进评估协议和引入遮挡感知极化机制,提升遮挡区域的预测性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.06769 2026-02-09 cs.LG 50%

Soft Forward-Backward Representations for Zero-shot Reinforcement Learning with General Utilities

用于具有通用效用的零样本强化学习的软前向-后向表示

Marco Bagatella, Thomas Rupf, Georg Martius, Andreas Krause

机构 * ETH Zurich, Zurich, Switzerland(苏黎世联邦理工学院) Max Planck Institute for Intelligent Systems, Tubingen, Germany(智能系统马克斯·普朗克研究所) University of Tubingen, Tubingen, Germany(图宾根大学)

专题命中 BEV与占用 :occupancy(abstract)

AI总结 本文提出了一种软前向-后向算法,用于解决具有通用效用的零样本强化学习问题,通过离线数据恢复随机策略并直接优化通用效用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.06433 2026-02-09 cs.CR cs.AR 50%

The Avatar Cache: Enabling On-Demand Security with Morphable Cache Architecture

Avatar缓存:通过可变形缓存架构实现按需安全

Anubhav Bhatla, Navneet Navneet, Moinuddin Qureshi, Biswabandan Panda

专题命中 BEV与占用 :occupancy(abstract)

AI总结 Avatar通过可变形缓存架构在传统LLC基础上实现按需安全,兼顾性能、功率和面积开销。

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 激光雷达 1 篇

2602.06380 2026-02-09 cs.RO cs.SY eess.SY 79%

A Consistency-Improved LiDAR-Inertial Bundle Adjustment

改进一致性 的 LiDAR-惯性束调整

Xinran Li, Shuaikang Zheng, Pengcheng Zheng, Xinyang Wang, Jiacheng Li, Zhitian Li, Xudong Zou

机构 * Aerospace Information Research Institute, Chinese Academy of Sciences(中国科学院航空信息研究所) University of Chinese Academy of Sciences(中国科学院大学) Aerospace Information Technology University(航空信息科技大学) QiLu Aerospace Information Research Institute, Chinese Academy of Sciences(齐鲁航空信息研究所) School of Electronic, Electrical and Communication Engineering(电子电气与通信工程学院)

专题命中 激光雷达 :LiDAR(title,abstract);分类 cs.RO

AI总结 本文提出改进一致性的 LiDAR-惯性束调整方法,通过定制参数化和估计器提升 SLAM 系统的准确性和可观测性。

详情

展开后加载摘要…

URL PDF HTML 收藏

6. 多传感器融合 1 篇

2511.14469 2026-02-09 cs.CV 57%

CompEvent: Complex-valued Event-RGB Fusion for Low-light Video Enhancement and Deblurring

CompEvent: 复数值事件-RGB融合用于低光照视频增强与去模糊

Mingchen Zhong, Xin Lu, Dong Li, Senyan Xu, Ruixuan Jiang, Xueyang Fu, Baocai Yin

专题命中 多传感器融合 :autonomous driving(abstract);分类 cs.CV

AI总结 CompEvent通过复数神经网络实现事件数据与RGB帧的全过程融合,提升低光照视频去模糊效果。

详情

展开后加载摘要…

URL PDF HTML 收藏