arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

自动驾驶

自动驾驶感知、规划、BEV、占用预测、激光雷达和仿真评测。

2026-02-12 至 2026-02-12 共收录 22 信号源:cs.RO, cs.CV, eess.IV, cs.AI

1. 感知 7 篇

2602.10660 2026-02-12 cs.CV 85%

AurigaNet: A Real-Time Multi-Task Network for Enhanced Urban Driving Perception

AurigaNet: 一种用于增强城市驾驶感知的实时多任务网络

Kiarash Ghasemzadeh, Sedigheh Dehghani

机构 * University of Alberta(阿尔伯塔大学) Shahid Beheshti University(沙希德·贝赫什提大学)

专题命中 感知 :driving perception(title,abstract);autonomous driving(abstract);self-driving(abstract);分类 cs.CV

AI总结 AurigaNet通过端到端实例分割提升自动驾驶感知性能,实现高精度的可驾驶区域分割、车道检测和目标识别。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.10771 2026-02-12 cs.CV cs.RO 81%

From Steering to Pedalling: Do Autonomous Driving VLMs Generalize to Cyclist-Assistive Spatial Perception and Planning?

从转向到踩踏:自动驾驶VLMs是否能泛化到骑行辅助的空间感知与规划?

Krishna Kanth Nakka, Vedasri Nakka

专题命中 感知 :autonomous driving(title,abstract);分类 cs.RO、cs.CV

AI总结 本文提出CyclingVQA基准,评估自动驾驶VLMs在骑行辅助场景中的感知与推理能力,发现现有模型在骑行特定交通提示理解上存在不足,需进一步改进。

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.10160 2026-02-12 cs.CV cs.AI 81%

AD$^2$: Analysis and Detection of Adversarial Threats in Visual Perception for End-to-End Autonomous Driving Systems

AD$^2$:面向端到端自动驾驶系统的视觉感知中对抗威胁的分析与检测

Ishan Sahu, Somnath Hazra, Somak Aditya, Soumyajit Dey

机构 * Indian Institute of Technology Kharagpur(印度理工学院卡里格普尔分校) TCS Research, India(印度塔塔咨询公司研究)

专题命中 感知 :autonomous driving(title,abstract);分类 cs.CV、cs.AI

AI总结 AD$^2$通过基于注意力机制的轻量级模型,检测端到端自动驾驶系统中由物理模糊、电磁干扰和数字攻击引起的对抗威胁,提升系统安全性。

Comments Accepted to WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.10943 2026-02-12 cs.CV cs.RO 62%

Towards Learning a Generalizable 3D Scene Representation from 2D Observations

从2D观测向量学习通用的3D场景表示

Martin Gromniak, Jan-Gerrit Habekost, Sebastian Kamp, Sven Magg, Stefan Wermter

机构 * University of Hamburg - Department of Informatics(汉堡大学信息学院) ZAL Center of Applied Aeronautical Research(应用航空研究中心) Hamburger Informatik Technologie-Center e.V. (HITeC)(汉堡信息科技中心(HITeC))

专题命中 感知 :occupancy(abstract);分类 cs.RO、cs.CV

AI总结 本文提出了一种通用神经辐射场方法,通过第一人称机器人观测学习3D场景表示,实现了对未见场景的泛化能力,并在真实场景中验证了其高精度的3D重建性能。

Comments Paper accepted at ESANN 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.09740 2026-02-12 cs.CV 57%

Robust Vision Systems for Connected and Autonomous Vehicles: Security Challenges and Attack Vectors

连接与自动驾驶车辆的稳健视觉系统:安全挑战与攻击向量

Sandeep Gupta, Roberto Passerone

机构 * Centre for Secure Information Technologies (CSIT), Queen's University Belfast, UK(安全信息科技中心(CSIT),女王大学贝尔法斯特,英国) Department of Information Engineering and Computer Science, University of Trento, Italy(信息工程与计算机科学系,特伦托大学,意大利)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

AI总结 本文研究了连接与自动驾驶车辆中视觉系统的鲁棒性,分析了关键传感器和组件,识别潜在攻击面并评估其对CIA原则的影响。

Comments Submitted to IEEE Transactions on Intelligent Vehicles

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21769 2026-02-12 cs.CV 57%

H2OFlow: Grounding Human-Object Affordances with 3D Generative Models and Dense Diffused Flows

H2OFlow: 通过3D生成模型和密集扩散流接地人类-物体 affordances

Harry Zhang, Luca Carlone

机构 * MIT(麻省理工学院)

专题命中 感知 :occupancy(abstract);分类 cs.CV

AI总结 H2OFlow通过3D生成模型和密集扩散流学习人类-物体交互的3D affordances,无需人工标注,有效泛化至现实物体。

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.06217 2026-02-12 quant-ph physics.optics 50%

Chernoff Information Bottleneck for Covert Quantum Target Sensing

Chernoff 信息瓶颈用于隐蔽量子目标感知

Giuseppe Ortolano, Ivano Ruo-Berchera, Leonardo Banchi

专题命中 感知 :LiDAR(abstract)

AI总结 本文提出基于Chernoff信息瓶颈原理的框架,展示纠缠光子探测器在隐蔽目标感知中的优势,为提升LiDAR和雷达的隐蔽性能提供新方法。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 规划控制 3 篇

2602.10365 2026-02-12 cs.RO math.OC 79%

Solving Geodesic Equations with Composite Bernstein Polynomials for Trajectory Planning

利用复合伯恩斯坦多项式求解测地方程用于轨迹规划

Nick Gorman, Gage MacLin, Maxwell Hammond, Venanzio Cichella

机构 * Department of Mechanical Engineering(机械工程系) University of Iowa(爱荷华大学)

专题命中 规划控制 :trajectory planning(title,abstract);分类 cs.RO

AI总结 本文提出基于复合伯恩斯坦多项式的轨迹规划方法,用于在复杂环境中生成连续、动态可行的轨迹,适用于多种自主系统。

Comments Accepted for the 2026 IEEE Aerospace Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.00437 2026-02-12 cs.CV 79%

ADGaussian: Generalizable Gaussian Splatting for Autonomous Driving via Multi-modal Joint Learning

ADGaussian:通过多模态联合学习实现自动驾驶的通用高斯点云重建

Qi Song, Chenghong Li, Haotong Lin, Sida Peng, Rui Huang

机构 * School of Science and Engineering, The Chinese University of Hong Kong (Shenzhen)(中国香港中文大学(深圳)科学与工程学院) Zhejiang University(浙江大学)

专题命中 规划控制 :autonomous driving(title);LiDAR(abstract);分类 cs.CV

AI总结 ADGaussian通过多模态联合学习实现自动驾驶场景的高斯点云重建,提升零样本泛化能力。

Comments The paper is accepted by ICRA 2026 and the project page can be found at https://maggiesong7.github.io/research/ADGaussian/

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11005 2026-02-12 cs.CV 57%

Interpretable Vision Transformers in Monocular Depth Estimation via SVDA

通过SVDA实现单目深度估计中的可解释性视觉变换器

Vasileios Arampatzakis, George Pavlidis, Nikolaos Mitianoudis, Nikos Papamarkos

机构 * Dept. Electrical and Computer Engineering(电子与计算机工程系) Democritus University of Thrace(德米特里乌斯大学) Athena Research Center(雅典研究中心) University Campus at Kimmeria(基米里亚大学校园)

专题命中 规划控制 :autonomous driving(abstract);分类 cs.CV

AI总结 SVDA通过引入谱结构化的注意力机制,提升了单目深度估计的可解释性,同时保持了预测精度并降低了计算开销。

Comments 8 pages, 2 figures, submitted to CVPR Conference 2026

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 端到端驾驶 2 篇

2602.10884 2026-02-12 cs.CV 83%

ResWorld: Temporal Residual World Model for End-to-End Autonomous Driving

ResWorld: 用于端到端自动驾驶的时序残差世界模型

Jinqing Zhang, Zehua Fu, Zelin Xu, Wenying Dai, Qingjie Liu, Yunhong Wang

机构 * State Key Laboratory of Virtual Reality Technology and Systems, Beihang University, Beijing, China(虚拟现实技术与系统国家重点实验室,北京航空航天大学) Zhongguancun Laboratory, Beijing, China(中关村实验室) Beijing Jingwei Hirain Technologies Co., Inc.(北京京wei Hirain科技有限公司) Hangzhou Innovation Institute, Beihang University, Hangzhou, China(杭州创新研究院,北京航空航天大学)

专题命中 端到端驾驶 :autonomous driving(title,abstract);BEV(abstract);分类 cs.CV

AI总结 ResWorld通过时序残差世界模型和未来引导轨迹细化模块,提升端到端自动驾驶的规划性能。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.10458 2026-02-12 cs.AI cs.LG 79%

Found-RL: foundation model-enhanced reinforcement learning for autonomous driving

Found-RL: 基于基础模型的强化学习用于自动驾驶

Yansong Qu, Zihao Sheng, Zilin Huang, Jiancong Chen, Yuhao Luo, Tianyi Wang, Yiheng Feng, Samuel Labi, Sikai Chen

专题命中 端到端驾驶 :autonomous driving(title,abstract);分类 cs.AI

AI总结 Found-RL通过异步批量推理和多样化监督机制,提升自动驾驶中强化学习的效率与实时性。

Comments 39 pages

详情

展开后加载摘要…

URL PDF HTML 收藏

4. BEV与占用 2 篇

2602.10738 2026-02-12 physics.hist-ph cond-mat.stat-mech quant-ph 50%

Between equilibrium and fluctuation: Einstein's heuristic argument and Boltzmann's principle

在平衡与波动之间:爱因斯坦的启发式论证与玻尔兹曼原理

Enric Pérez, Antonio Gil

专题命中 BEV与占用 :occupancy(abstract)

AI总结 本文探讨爱因斯坦1905年光量子论证的模糊性及其与玻尔兹曼原理的关系,指出其局限性在于频率而非占据数。

Comments 36 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.10262 2026-02-12 cs.DC cs.AR 50%

Execution-Centric Characterization of FP8 Matrix Cores, Asynchronous Execution, and Structured Sparsity on AMD MI300A

面向执行的FP8矩阵核心、异步执行与结构稀疏性在AMD MI300A上的特性分析

Aaron Jarmusch, Connor Vitz, Sunita Chandrasekaran

专题命中 BEV与占用 :occupancy(abstract)

AI总结 本文通过微观基准测试分析了AMD MI300A上的FP8矩阵执行、ACE并发性和结构稀疏性,揭示了其执行特性及对HPC和HPC-AI工作负载性能的影响。

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 激光雷达 6 篇

2602.10492 2026-02-12 cs.CV cs.RO 82%

End-to-End LiDAR optimization for 3D point cloud registration

端到端激光雷达优化用于3D点云配准

Siddhant Katyan, Marc-André Gardner, Jean-François Lalonde

机构 * Université Laval(拉瓦尔大学) Bentley Systems(贝恩特系统)

专题命中 激光雷达 :LiDAR(title,abstract);分类 cs.RO、cs.CV

AI总结 本文提出端到端自适应激光雷达框架,通过动态调整参数优化点云配准,提升精度与效率,并在CARLA模拟中验证其优于固定参数方法的性能。

Comments 36th British Machine Vision Conference 2025, {BMVC} 2025, Sheffield, UK, November 24-27, 2025. Project page: https://lvsn.github.io/e2e-lidar-registration/

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.10837 2026-02-12 eess.IV 79%

FPGA Implementation of Sketched LiDAR for a 192 x 128 SPAD Image Sensor

FPGA实现基于多项式样条函数的压缩LiDAR用于192 x 128 SPAD图像传感器

Zhenya Zang, Mike Davies, Istvan Gyongy

专题命中 激光雷达 :LiDAR(title,abstract);分类 eess.IV

AI总结 本研究提出了一种基于多项式样条函数的FPGA实现方法,用于压缩192 x 128 SPAD图像传感器的数据,实现高保真度的在线深度重建,解决SPAD阵列的时间戳传输瓶颈问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.18301 2026-02-12 cs.CV 79%

Contextual Range-View Projection for 3D LiDAR Point Clouds

基于上下文的视距投影用于3D激光雷达点云

Seyedali Mousavi, Seyedhamidreza Mousavi, Masoud Daneshtalab

机构 * School of Innovation, Design and Engineering(创新、设计与工程学院) Division of Intelligent Future Technologies(智能未来技术系) Mälardalen University(马尔默大学)

专题命中 激光雷达 :LiDAR(title,abstract);分类 cs.CV

AI总结 本文提出CAP和CWAP两种机制,通过结合实例中心和类别标签的上下文信息,改进3D激光雷达点云的视距投影,提升实例点保留率和目标类别性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11082 2026-02-12 cs.RO eess.SP 57%

Digging for Data: Experiments in Rock Pile Characterization Using Only Proprioceptive Sensing in Excavation

挖掘数据:仅使用本体感觉传感在挖掘中进行岩石堆 characterization 的实验

Unal Artan, Martin Magnusson, Joshua A. Marshall

机构 * Center for Applied Autonomous Sensor Systems, Örebro University(应用自主传感系统中心,奥雷布罗大学) Ingenuity Labs Research Institute, Smith Engineering, Queen’s University(创新实验室研究机构,史密斯工程,女王大学)

专题命中 激光雷达 :LiDAR(abstract);分类 cs.RO

AI总结 本研究提出了一种仅使用本体感觉传感数据来估计破碎岩石堆相对颗粒大小的方法,通过现场实验验证了该方法的有效性,并与视觉分析和筛分方法进行了对比。

Comments Accepted for publication in the IEEE Transactions on Field Robotics

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.10871 2026-02-12 cs.HC cs.CV 57%

Viewpoint Recommendation for Point Cloud Labeling through Interaction Cost Modeling

通过交互成本建模实现点云标注的视角推荐

Yu Zhang, Xinyi Zhao, Chongke Bi, Siming Chen

机构 * School of Data Science, Fudan University(复旦大学数据科学学院) Department of Computer Science, University of Oxford(牛津大学计算机科学系) College of Intelligence and Computing, Tianjin University(天津大学智能与计算学院) Shanghai Key Laboratory of Data Science, Shanghai(上海数据科学重点实验室)

专题命中 激光雷达 :autonomous driving(abstract);分类 cs.CV

AI总结 本文提出通过交互成本建模减少点云标注时间的方法,利用Fitts定律优化视角推荐,提升标注效率。

Comments Accepted to IEEE TVCG

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11026 2026-02-12 cs.HC 50%

Normalized Surveillance in the Datafied Car: How Autonomous Vehicle Users Rationalize Privacy Trade-offs

数据化汽车中的规范化监控:自动驾驶车辆用户如何合理化隐私权衡

Yehuda Perry, Tawfiq Ammari

专题命中 激光雷达 :LiDAR(abstract)

AI总结 研究探讨自动驾驶车辆用户如何通过比较现有数字平台来合理化隐私权衡,并提出治理措施以规范数据提取。

详情

展开后加载摘要…

URL PDF HTML 收藏

6. 仿真评测 1 篇

2511.03220 2026-02-12 eess.SP 50%

Multimodal-Wireless: A Large-Scale Dataset for Sensing and Communication

多模态无线:一种大规模数据集用于传感与通信

Tianhao Mao, Le Liang, Jie Yang, Hao Ye, Shi Jin, Geoffrey Ye Li

专题命中 仿真评测 :LiDAR(abstract)

AI总结 本文提出多模态无线数据集,用于多模态传感与通信研究,包含高分辨率CSI与多种传感器数据,支持通信与协同感知应用。

详情

展开后加载摘要…

URL PDF HTML 收藏

7. 其他自动驾驶 1 篇

2511.20022 2026-02-12 cs.CV cs.AI 81%

WaymoQA: A Multi-View Visual Question Answering Dataset for Safety-Critical Reasoning in Autonomous Driving

WaymoQA: 一个用于自动驾驶安全关键推理的多视角视觉问答数据集

Seungjun Yu, Seonho Lee, Namho Kim, Jaeyo Shin, Junsung Park, Wonjeong Ryu, Raehyuk Jung, Hyunjung Shim

机构 * Hanyang University(翰阳大学)

专题命中 其他自动驾驶 :autonomous driving(title,abstract);分类 cs.CV、cs.AI

AI总结 WaymoQA通过多视角输入提升自动驾驶的安全关键推理能力,提供35,000个标注问题-答案对,显著增强大语言模型的推理性能。

详情

展开后加载摘要…

URL PDF HTML 收藏