arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

自动驾驶

自动驾驶感知、规划、BEV、占用预测、激光雷达和仿真评测。

共收录 6072 信号源:cs.RO, cs.CV, eess.IV, cs.AI

1. 感知 6072 篇

2604.04887 2026-04-07 cs.CV 70%

HorizonWeaver: Generalizable Multi-Level Semantic Editing for Driving Scenes

HorizonWeaver: 通用多级语义编辑用于驾驶场景

Mauricio Soroco, Francesco Pittaluga, Zaid Tasneem, Abhishek Aich, Bingbing Zhuang, Wuyang Chen, Manmohan Chandraker, Ziyu Jiang

机构 * Simon Fraser University(西蒙菲莎大学) NEC Labs America(NEC美国实验室)

专题命中 感知 :autonomous driving(abstract);BEV(abstract);分类 cs.CV

AI总结 本文提出HorizonWeaver,通过多级粒度、丰富高层语义和普遍领域转移解决驾驶场景编辑难题,构建大规模数据集并改进模型与训练方法,实现复杂驾驶场景的高质量指令驱动编辑。

Comments CVPR Findings 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.14787 2026-04-07 cs.CV 70%

Polarimetric Imaging for Perception

极化成像用于感知

Michael Baltaxe, Tomer Pe'er, Dan Levi

机构 * General Motors(通用汽车)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 本文研究极化成像在自动驾驶中的应用,利用RGB极化相机提升单目深度估计和自由空间检测性能,提出新数据集支持极化信息感知算法开发。

Journal ref British Machine Vision Conference (BMVC) 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16483 2026-04-02 cs.CV 70%

Octree Diffusion for Semantic Scene Generation and Completion

八叉树扩散用于语义场景生成与补全

Xujia Zhang, Brendan Crowe, Christoffer Heckman

机构 * University of Colorado Boulder(科罗拉多大学博尔德分校) Autonomous Robotics and Perception Group(自主机器人与感知实验室)

专题命中 感知 :LiDAR(abstract);occupancy(abstract);分类 cs.CV

AI总结 本文提出Octree Latent Semantic Diffusion框架,通过八叉树隐式表示实现跨域的3D语义场景补全、扩展和生成,结合结构扩散与语义扩散生成语义标签,无需重训练即可完成单扫描的补全与零样本泛化。

Comments Accepted to ICRA 2026. Revised version with updated paragraphs

Journal ref Proceedings of the 2026 IEEE International Conference on Robotics and Automation (ICRA)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.02413 2026-04-01 cs.CV 70%

TruckDrive: Long-Range Autonomous Highway Driving Dataset

TruckDrive:长距离自动驾驶高速公路数据集

Filippo Ghilotti, Edoardo Palladin, Samuel Brucker, Adam Sigal, Mario Bijelic, Felix Heide

机构 * Torc Robotics Princeton University(普林斯顿大学)

专题命中 感知 :autonomous driving(abstract);driving perception(abstract);分类 cs.CV

AI总结 本文提出TruckDrive数据集,用于解决重卡车高速公路自动驾驶中长距离感知的挑战,通过多模态传感器采集47.5万样本,支持20秒序列中2D检测1000米和3D检测400米的感知基准测试。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.27995 2026-03-31 cs.CV 70%

UniDA3D: A Unified Domain-Adaptive Framework for Multi-View 3D Object Detection

UniDA3D:一种统一的多视图3D目标检测领域自适应框架

Hongjing Wu, Cheng Chi, Jinlin Wu, Yanzhao Su, Zhen Lei, Wenqi Ren

机构 * Shenzhen Campus of Sun Yat-sen University(中山大学深圳校区) The State Key Laboratory of Blockchain and Data Security, Zhejiang University(浙江大学区块链与数据安全全国重点实验室) Beijing Academy of Artificial Intelligence(北京人工智能研究院) National Laboratory of Pattern Recognition, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所模式识别国家重点实验室)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 UniDA3D通过统一领域自适应方法提升多视图3D目标检测在复杂环境下的鲁棒性,采用新颖的查询引导领域差异缓解模块和动态伪标签训练策略,实现跨多领域的一致性学习。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11306 2026-03-26 cs.RO 70%

Rotor-Failure-Aware Quadrotors Flight in Unknown Environments

具有旋转故障意识的四旋翼在未知环境中的飞行

Xiaobin Zhou, Miao Wang, Chengao Li, Can Cui, Ruibin Zhang, Yongchao Wang, Chao Xu, Fei Gao

机构 * School of Robotics and Automation, Nanjing University(南京大学机器人与自动化学院) Institute of Cyber-Systems and Control, College of Control Science and Engineering, Zhejiang University(浙江大学控制系统研究所) Department of Aeronautical and Aviation Engineering, The Hong Kong Polytechnic University(香港理工大学航空工程系) School of Aeronautic Science and Engineering, Beihang University(北航航空科学与工程学院)

专题命中 感知 :LiDAR(abstract);trajectory planning(abstract);分类 cs.RO

AI总结 本文提出一种具有旋转故障意识的四旋翼导航系统,通过复合故障检测与诊断非线性模型预测控制器和LiDAR平台,在未知复杂环境中实现旋转故障的快速检测与飞行稳定。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.22781 2026-03-25 cs.CV 70%

Typography-Based Monocular Distance Estimation Framework for Vehicle Safety Systems

基于字体的单目距离估计框架用于车辆安全系统

Manognya Lokesh Reddy, Zheng Liu

机构 * Department of Computer and Information Science, University of Michigan-Dearborn(密歇根大学迪尔伯恩分校计算机与信息科学系) Department of Industrial and Manufacturing Systems Engineering, University of Michigan-Dearborn(密歇根大学迪尔伯恩分校工业与制造系统工程系)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 本文提出基于车牌标准化字体的单目距离估计方法,利用字符高度和针孔相机模型计算距离,结合多种鲁棒性增强技术,实现高精度实时距离估计。

Comments 25 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00385 2026-03-24 cs.CV 70%

EZ-SP: Fast and Lightweight Superpoint-Based 3D Segmentation

EZ-SP:快速且轻量的基于超点的3D分割

Louis Geist, Loic Landrieu, Damien Robert

机构 * LIGM, ENPC, IP Paris, Univ Gustave Eiffel, CNRS(LIGM,ENPC,IP巴黎,巴黎理工大学,CNRS) DM3L, University of Zurich(DM3L,苏黎世大学)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 本文提出了一种高效的基于超点的3D语义分割方法,通过全GPU分区算法实现13倍更快的处理速度,结合轻量级分类器,达到2MB显存占用,支持实时推理,并在三个领域验证了其准确性。

Comments Accepted at ICRA 2026. Camera-ready version with Appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19235 2026-03-24 cs.CV 70%

IDSplat: Instance-Decomposed 3D Gaussian Splatting for Driving Scenes

IDSplat:实例分解的3D高斯点云法用于驾驶场景

Carl Lindström, Mahan Rafidashti, Maryam Fatemi, Lars Hammarstrand, Martin R. Oswald, Lennart Svensson

机构 * Zenseact Chalmers University of Technology(瑞典查尔姆斯理工大学) University of Amsterdam(阿姆斯特丹大学)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 IDSplat通过实例分解和可学习运动轨迹重建动态驾驶场景,无需人工标注,实现场景分离与泛化。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02245 2026-03-23 cs.RO 70%

CoInfra: A Large-Scale Cooperative Infrastructure Perception System and Dataset for Vehicle-Infrastructure Cooperation in Adverse Weather

CoInfra:一种大规模协作基础设施感知系统及数据集,用于恶劣天气下的车-基础设施协作

Minghao Ning, Yufeng Yang, Keqi Shu, Shucheng Huang, Jiaming Zhong, Maryam Salehi, Mahdi Rahmani, Jiaming Guo, Yukun Lu, Chen Sun, Aladdin Saleh, Ehsan Hashemi, Amir Khajepour

机构 * Department of Mechanical and Mechatronics Engineering, University of Waterloo(滑铁卢大学机械与机电工程系) Department of Mechanical Engineering, University of New Brunswick(新不伦瑞克大学机械工程系) Department of Data and Systems Engineering, University of Hong Kong(香港大学数据与系统工程系) Technology Partnerships and Innovations, Rogers Communications, Canada Inc.(罗杰斯通讯公司技术创新部) Department of Mechanical Engineering, University of Alberta(阿尔伯塔大学机械工程系)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.RO

AI总结 本文提出CoInfra系统,通过14个路侧传感器节点和5G网络实现多节点协同感知,涵盖8个节点的环形城市交叉口在四种天气条件下的数据,验证基础设施感知在复杂交通场景中的安全提升作用。

Comments This paper has been submitted to the Transportation Research Part C: Emerging Technologies for review

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16119 2026-03-17 cs.CV 70%

RadarGaussianDet3D: Gaussian Representation-based Real-time 3D Object Detection with 4D Automotive Radars

RadarGaussianDet3D: 基于高斯表示的实时3D目标检测与4D汽车雷达

Weiyi Xiong, Bing Zhu, Zewei Zheng

专题命中 感知 :autonomous driving(abstract);BEV(abstract);分类 cs.CV

AI总结 本文提出RadarGaussianDet3D,通过高斯表示提升4D雷达的3D目标检测性能,采用点高斯编码器和3D高斯点划技术生成密集特征图,结合新的高斯损失函数实现更高效的检测与推理。

Comments Accepted by IEEE Robotics and Automation Letters

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.15423 2026-03-13 cs.RO cs.SY eess.SY 70%

Online Slip Detection and Friction Coefficient Estimation for Autonomous Racing

自动驾驶赛车中的在线滑动检测与摩擦系数估计

Christopher Oeltjen, Carson Sobolewski, Saleh Faghfoorian, Lorant Domokos, Giancarlo Vidal, Sriram Yerramsetty, Ivan Ruchkin

机构 * Trustworthy Engineered Autonomy (TEA) Lab, Department of Electrical and Computer Engineering, University of Florida(可信工程自主性实验室,电气与计算机工程系,佛罗里达大学)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.RO

AI总结 本文提出了一种基于IMU和LiDAR数据的轻量级方法,用于自动驾驶赛车中的在线滑动检测和摩擦系数估计,无需复杂模型或训练数据。

Comments Equal contribution by the first three authors

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.07486 2026-03-10 cs.CV 70%

Multi-Modal Decouple and Recouple Network for Robust 3D Object Detection

多模态解耦与耦合网络用于抗干扰的3D目标检测

Rui Ding, Zhaonian Kuang, Yuzhe Ji, Meng Yang, Xinhu Zheng, Gang Hua

机构 * State Key Laboratory of Human-Machine Hybrid Augmented Intelligence, Institute of Artificial Intelligence and Robotics, Xi’an Jiaotong University(人机混合增强智能国家重点实验室,人工智能与机器人研究院,西安交通大学) Intelligent Transportation Thrust of the Systems Hub, The Hong Kong University of Science and Technology (Guangzhou)(系统枢纽智能交通方向,香港科技大学(广州)) Multimodal Experiences Research Lab, Dolby Laboratories(多模态体验研究实验室,Dolby实验室)

专题命中 感知 :BEV(abstract);LiDAR(abstract);分类 cs.CV

AI总结 本文提出多模态解耦与耦合网络,通过分离和重新耦合不同模态特征以提高在数据损坏下的3D目标检测鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.02481 2026-03-04 cs.CV 70%

ModalPatch: A Plug-and-Play Module for Robust Multi-Modal 3D Object Detection under Modality Drop

ModalPatch: 一种插拔式模块,用于在模态缺失情况下鲁棒的多模态3D目标检测

Shuangzhi Li, Lei Ma, Xingyu Li

机构 * University of Alberta, Canada(阿尔伯塔大学) The University of Tokyo, Japan(东京大学)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 ModalPatch是一种插拔式模块,通过利用传感器数据的时间特性实现感知连续性,提升多模态3D目标检测在模态缺失情况下的鲁棒性和准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.23871 2026-03-02 cs.CV cs.LG 70%

Bandwidth-adaptive Cloud-Assisted 360-Degree 3D Perception for Autonomous Vehicles

带宽自适应的云辅助360度3D感知用于自动驾驶车辆

Faisal Hawladera, Rui Meireles, Gamal Elghazaly, Ana Aguiar, Raphaël Frank

机构 * Interdisciplinary Centre for Security, Reliability, and Trust (SnT), University of Luxembourg, L-1855, Luxembourg(安全、可靠性与信任跨学科研究中心(SnT),卢森堡大学) Computer Science Department, Vassar College, Poughkeepsie, NY 12604, USA(计算机科学系,瓦萨学院)

专题命中 感知 :autonomous driving(abstract);BEV(abstract);分类 cs.CV

AI总结 本文提出一种带宽自适应的云辅助3D感知方法,通过动态分割计算任务和特征压缩,降低延迟并提升检测精度,适用于复杂城市环境中的自动驾驶。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20632 2026-02-25 cs.CV 70%

Boosting Instance Awareness via Cross-View Correlation with 4D Radar and Camera for 3D Object Detection

通过4D雷达和相机的跨视图相关性提升实例意识

Xiaokai Bai, Lianqing Zheng, Si-Yuan Cao, Xiaohan Zhang, Zhe Wu, Beinan Yu, Fang Wang, Jie Bai, Hui-Liang Shen

机构 * College of Information Science and Electronic Engineering, Zhejiang University(浙江大学信息科学与电子工程学院) School of Automotive Studies, Tongji University(同济大学汽车学院) College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院) Jinhua Institute of Zhejiang University(浙江大学金华研究院) School of Information and Electrical Engineering, Hangzhou City University(杭州城市学院信息与电气工程学院) Hangzhou City University Binjiang Innovation Center(杭州城市学院滨江创新中心)

专题命中 感知 :autonomous driving(abstract);BEV(abstract);分类 cs.CV

AI总结 SIFormer通过结合4D雷达和相机的跨视图相关性,提升3D目标检测中的实例意识,结合两种融合范式的优点,提高检测精度。

Comments 14 pages, 10 figures, 13 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06685 2026-02-24 cs.CV 70%

MOGS: Monocular Object-guided Gaussian Splatting in Large Scenes

MOGS:在大场景中使用单目物体引导的高斯散射

Shengkai Zhang, Yuhe Liu, Jianhua He, Xuedou Xiao, Mozi Chen, Kezhong Liu

机构 * State Key Laboratory of Maritime Technology and Safety, Wuhan University of Technology(船舶技术与安全国家重点实验室,武汉理工大学) University of Essex(埃塞克斯大学)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 MOGS通过单目视觉-惯性传感器实现大场景中高斯散射的物体引导密集深度生成,显著降低训练时间和内存消耗,同时保持高质量渲染效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22890 2026-02-16 cs.CV cs.CR 70%

CP-uniGuard: A Unified, Probability-Agnostic, and Adaptive Framework for Malicious Agent Detection and Defense in Multi-Agent Embodied Perception Systems

CP-uniGuard: 多智能体具身体验系统中恶意代理检测与防御的统一、概率无关和自适应框架

Senkang Hu, Yihang Tao, Guowen Xu, Xinyuan Qian, Yiqin Deng, Xianhao Chen, Sam Tak Wu Kwong, Yuguang Fang

机构 * Hong Kong JC STEM Lab of Smart City and Department of Computer Science, City University of Hong Kong(香港JC STEM实验室及城市大学计算机科学系) School of Computer Science and Engineering, University of Electronic Science and Technology of China(电子科技大学计算机科学与工程学院) Department of Electrical and Electronic Engineering, The University of Hong Kong(香港大学电子与电气工程系) School of Data Science, Lingnan University(岭南大学数据科学学院)

专题命中 感知 :autonomous driving(abstract);BEV(abstract);分类 cs.CV

AI总结 CP-uniGuard通过概率无关的样本共识和自适应阈值,实现多智能体系统中恶意代理的检测与防御。

Comments Accepted by IEEE Transactions on Mobile Computing (TMC)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.06406 2026-02-09 cs.CV 70%

Point Virtual Transformer

点虚拟变换器

Veerain Sood, Bnalin, Gaurav Pandey

机构 * Texas A \& M University Email Engineering Technology \& Industrial Distribution Texas A \& M University Texas, USA

专题命中 感知 :BEV(abstract);LiDAR(abstract);分类 cs.CV

AI总结 PointViT通过融合真实和虚拟点,提升远距离3D目标检测性能,实现91.16%的3D AP和95.94%的BEV AP。

Comments 8 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.05538 2026-02-09 cs.CV 70%

A Comparative Study of 3D Person Detection: Sensor Modalities and Robustness in Diverse Indoor and Outdoor Environments

三维人物检测的比较研究:传感器模态与在多样室内和室外环境中的鲁棒性

Malaz Tamim, Andrea Matic-Flierl, Karsten Roscher

机构 * Fraunhofer Institute for Cognitive Systems IKS(弗劳恩霍夫认知系统研究所)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 本文比较了三种三维人物检测方法,发现融合模型在复杂场景中表现最佳,但对传感器错位仍敏感。

Comments Accepted for VISAPP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.14325 2026-02-05 cs.CV 70%

Unlocking Past Information: Temporal Embeddings in Cooperative Bird's Eye View Prediction

解锁过去信息:合作鸟瞰图预测中的时间嵌入

Dominik Rößle, Jeremias Gerner, Klaus Bogenberger, Daniel Cremers, Stefanie Schmidtner, Torsten Schön

机构 * Department of Computer Science and AImotion Bavaria, Technische Hochschule Ingolstadt(计算机科学系和AImotion巴伐利亚,因戈尔施塔特技术大学) Department of Electrical Engineering and AImotion Bavaria, Technische Hochschule Ingolstadt(电气工程系和AImotion巴伐利亚,因戈尔施塔特技术大学) School of Engineering and Design, Technical University of Munich(工程与设计学院,慕尼黑技术大学) School of Computation, Information and Technology, Technical University of Munich(计算、信息与技术学院,慕尼黑技术大学)

专题命中 感知 :autonomous driving(abstract);BEV(abstract);分类 cs.CV

AI总结 TempCoBEV通过整合历史信息提升合作鸟瞰图预测的准确性和可靠性,尤其在通信故障情况下表现显著。

Comments Copyright 2024 IEEE. This is the accepted version of the paper. In 2024 IEEE Intelligent Vehicles Symposium (IV), pp. 2220-2225. Official paper available at https://doi.org/10.1109/IV55156.2024.10588608

Journal ref IEEE Intelligent Vehicles Symposium (IV), pp. 2220-2225, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02858 2026-02-04 cs.RO cs.LG cs.MA cs.NI cs.SY eess.SY 70%

IMAGINE: Intelligent Multi-Agent Godot-based Indoor Networked Exploration

IMAGINE: 基于 Godot 的智能多智能体室内网络探索

Tiago Leite, Maria Conceição, António Grilo

机构 * Institute for Systems and Robotics(系统研究所)

专题命中 感知 :LiDAR(abstract);occupancy(abstract);分类 cs.RO

AI总结 本文提出基于Godot的多智能体强化学习方法,用于解决室内未知环境中的自主协作探索问题,通过高保真模拟和课程学习提升训练效率与鲁棒性。

Comments 12 pages, submitted to a journal

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03906 2026-01-27 cs.CV 70%

From Filters to VLMs: Benchmarking Defogging Methods through Object Detection and Segmentation Performance

从滤波器到视觉语言模型:通过目标检测和分割性能评估去雾方法

Ardalan Aryashad, Parsa Razmara, Amin Mahjoub, Seyedarmin Azizi, Mahdi Salmani, Arad Firouzkouhi

机构 * University of Southern California(南加州大学)

专题命中 感知 :autonomous driving(abstract);driving perception(abstract);分类 cs.CV

AI总结 本文通过目标检测和分割性能评估,探讨了去雾方法在真实与合成环境中的有效性,揭示了视觉语言模型在恶劣天气下的应用潜力。

Comments Accepted at WACV 2026 Proceedings (Oral), 5th Workshop on Image, Video, and Audio Quality Assessment in Computer Vision, with a focus on VLM and Diffusion Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15545 2026-01-23 cs.CV 70%

Multi-View Projection for Unsupervised Domain Adaptation in 3D Semantic Segmentation

多视角投影用于无监督领域适应的3D语义分割

Andrew Caunes, Thierry Chateau, Vincent Fremont

机构 * Logiroad, Nantes, France(法国南特Logiroad) LS2N - Ecole Centrale de Nantes, France(法国南特LS2N-中央理工学院)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 本文提出了一种基于多视角投影的无监督领域适应方法,通过合成数据集训练2D分割模型并回投影生成3D标签,实现3D语义分割的领域适应和稀有类别分割。

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.10962 2026-01-19 cs.CV 70%

V2X-Radar: A Multi-modal Dataset with 4D Radar for Cooperative Perception

V2X-Radar: 一种包含4D雷达的多模态数据集用于协作感知

Lei Yang, Xinyu Zhang, Jun Li, Chen Wang, Jiaqi Ma, Zhiying Song, Tong Zhao, Ziying Song, Li Wang, Mo Zhou, Yang Shen, Kai Wu, Chen Lv

机构 * School of Vehicle and Mobility, Tsinghua University(车辆与移动性学院,清华大学) Nanyang Technological University(南洋理工大学) CUMTB(中国交通车辆技术研究所) University of California, Los Angeles(加州大学洛杉矶分校) Beijing Jiaotong University(北京交通大学) ByteDance(字节跳动)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 V2X-Radar是首个包含4D雷达的多模态数据集,用于提升自动驾驶中的协作感知能力。

Comments NeurIPS 2025 Spotlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04968 2026-01-09 cs.CV 70%

SparseLaneSTP: Leveraging Spatio-Temporal Priors with Sparse Transformers for 3D Lane Detection

SparseLaneSTP: 利用稀疏变换器结合时空先验进行3D车道检测

Maximilian Pittner, Joel Janai, Mario Faigle, Alexandru Paul Condurache

机构 * Bosch Mobility Solutions, Robert Bosch GmbH(博世移动解决方案,罗伯特·博世有限公司) Institute of Neuro- and Bioinformatics, University of Lübeck(神经与生物医学研究所,吕贝克大学) Institute for Signal Processing and System Theory, University of Stuttgart(信号处理与系统理论研究所,斯图加特大学)

专题命中 感知 :autonomous driving(abstract);BEV(abstract);分类 cs.CV

AI总结 SparseLaneSTP通过整合车道结构的几何属性和时间信息,利用稀疏变换器和时空注意力机制,提升3D车道检测的精度和一致性。

Comments Published at IEEE/CVF International Conference on Computer Vision (ICCV) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.03001 2026-01-07 cs.CV 70%

Towards Efficient 3D Object Detection for Vehicle-Infrastructure Collaboration via Risk-Intent Selection

为车辆-基础设施协作实现高效3D物体检测的险意选择

Li Wang, Boqi Li, Hang Chen, Xingjian Wu, Yichen Wang, Jiewen Tan, Xinyu Zhang, Huaping Liu

专题命中 感知 :autonomous driving(abstract);BEV(abstract);分类 cs.CV

AI总结 本文提出RiSe框架,通过风险意图选择实现车辆-基础设施协作中高效3D物体检测,减少通信量同时保持高检测精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.24922 2026-01-01 cs.CV 70%

Semi-Supervised Diversity-Aware Domain Adaptation for 3D Object detection

半监督多样性感知领域适应用于3D目标检测

Bartłomiej Olber, Jakub Winter, Paweł Wawrzyński, Andrii Gamalii, Daniel Górniak, Marcin Łojek, Robert Nowak, Krystian Radlak

机构 * Warsaw University of Technology(华沙技术大学) IDEAS NCBR

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 本文提出了一种基于神经激活模式的半监督领域适应方法,通过少量多样化样本标注提升3D目标检测在跨领域应用中的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.23215 2025-12-30 cs.CV 70%

AVOID: The Adverse Visual Conditions Dataset with Obstacles for Driving Scene Understanding

AVOID: 用于驾驶场景理解的障碍物条件数据集

Jongoh Jeong, Taek-Jin Song, Jong-Hwan Kim, Kuk-Jin Yoon

专题命中 感知 :self-driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 AVOID数据集通过模拟环境收集了多种恶劣天气条件下的道路障碍物,用于提升自动驾驶场景理解的实时障碍物检测能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.22503 2025-12-30 cs.CV 70%

SCAFusion: A Multimodal 3D Detection Framework for Small Object Detection in Lunar Surface Exploration

SCAFusion: 一种针对月球表面探索的小目标多模态3D检测框架

Xin Chen, Kang Luo, Yangyi Xiao, Hesheng Wang

机构 * Department of Automation, Key Laboratory of System Control and Information Processing of Ministry of Education, Key Laboratory of Marine Intelligent Equipment and System of Ministry of Education, Shanghai Engineering Research Center of Intelligent Control and Management, Shanghai Jiao Tong University(自动化系、教育部系统控制与信息处理重点实验室、教育部海洋智能装备与系统重点实验室、上海智能控制与管理工程研究中心、上海交通大学)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 SCAFusion提出了一种针对月球表面探索的小目标多模态3D检测框架,通过改进的特征对齐和坐标注意机制提升小目标检测性能。

详情

展开后加载摘要…

URL PDF HTML 收藏