arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

自动驾驶

自动驾驶感知、规划、BEV、占用预测、激光雷达和仿真评测。

共收录 6072 信号源:cs.RO, cs.CV, eess.IV, cs.AI

1. 感知 6072 篇

2512.19934 2025-12-24 cs.CV cs.AI cs.LG 62%

Vehicle-centric Perception via Multimodal Structured Pre-training

基于多模态结构预训练的车辆感知

Wentao Wu, Xiao Wang, Chenglong Li, Jin Tang, Bin Luo

机构 * Information Materials and Intelligent Sensing Laboratory of Anhui Province(安徽省信息材料与智能感知实验室) Anhui Provincial Key Laboratory of Multimodal Cognitive Computation(安徽省多模态认知计算重点实验室) the School of Artificial Intelligence, Anhui University(安徽大学人工智能学院) School of Computer Science and Technology, Anhui University(安徽大学计算机科学与技术学院) Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合性国家科学中心人工智能研究院)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

AI总结 本文提出VehicleMAE-V2,通过多模态结构先验知识提升车辆感知的预训练能力,采用SMM、CRM和SRM模块增强模型对车辆结构和语义的理解。

Comments Journal extension of VehicleMAE (AAAI 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.20034 2025-12-23 cs.CV cs.RO 62%

NeSLAM: Neural Implicit Mapping and Self-Supervised Feature Tracking With Depth Completion and Denoising

NeSLAM: 基于深度补全与去噪的神经隐式映射与自监督特征跟踪

Tianchen Deng, Yanbo Wang, Hongle Xie, Hesheng Wang, Jingchuan Wang, Danwei Wang, Weidong Chen

机构 * Institute of Medical Robotics and Department of Automation, Shanghai Jiao Tong University(上海交通大学医疗机器人研究所和自动化系) Key Laboratory of System Control and Information Processing, Ministry of Education, Shang hai 200240, China(教育部系统控制与信息处理重点实验室,上海200240,中国) School of Electrical and Electronic Engineering, Nanyang Technological University(南洋理工大学电子与电气工程学院)

专题命中 感知 :occupancy(abstract);分类 cs.RO、cs.CV

AI总结 NeSLAM通过深度补全与去噪网络和SDF表示,实现高精度的3D重建、鲁棒的相机跟踪和视点合成。

Journal ref IEEE Transactions on Automation Science and Engineering, vol. 22, pp. 12309-12321, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18082 2025-12-23 cs.CV cs.AI 62%

Uncertainty-Gated Region-Level Retrieval for Robust Semantic Segmentation

不确定性门控区域级检索用于鲁棒语义分割

Shreshth Rajan, Raymond Liu

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

AI总结 本文提出一种不确定性门控区域级检索机制,提升语义分割在域移位下的鲁棒性和精度,实现分割准确率提升11.3%且检索成本降低87.5%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.16025 2025-12-22 cs.LG cs.AI cs.RO 62%

Symbolic Imitation Learning: From Black-Box to Explainable Driving Policies

符号模仿学习:从黑箱到可解释的驾驶策略

Iman Sharifi, Mustafa Yildirim, Saber Fallah

机构 * Department of Science and Engineering, the George Washington University(科学与工程系,乔治·华盛顿大学) Autonomous Vehicles Lab, Department of Mechanical Engineering, University of Surrey(自主车辆实验室,机械工程系,萨里大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.RO、cs.AI

AI总结 本文提出符号模仿学习框架,利用归纳逻辑编程生成可解释的驾驶策略,通过实验证明其在提升透明度和性能方面的优势。

Comments 24 pages, 4 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.16123 2025-12-19 cs.CR cs.AI cs.CV 62%

Autoencoder-based Denoising Defense against Adversarial Attacks on Object Detection

基于自动编码器的对抗攻击防御方法用于目标检测

Min Geun Song, Gang Min Kim, Woonmin Kim, Yongsik Kim, Jeonghyun Sim, Sangbeom Park, Huy Kang Kim

机构 * School of Cybersecurity in Korea University(韩国大学网络安全学院)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

AI总结 本文提出基于自动编码器的去噪方法,用于防御目标检测中的对抗攻击,实验表明该方法能有效恢复检测性能,无需模型重训练。

Comments 7 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04870 2025-12-16 eess.IV cs.CV 62%

Multi-modal Uncertainty Robust Tree Cover Segmentation For High-Resolution Remote Sensing Images

多模态不确定性鲁棒树冠覆盖分割用于高分辨率遥感图像

Yuanyuan Gui, Wei Li, Yinjian Wang, Xiang-Gen Xia, Mauro Marty, Christian Ginzler, Zuyuan Wang

机构 * School of Information and Electronics, Beijing Institute of Technology(信息与电子学院,北京理工大学) National Key Laboratory of Science and Technology on Space-Born Intelligent Information Processing(空间智能信息处理国家重点实验室) Department of Electrical and Computer Engineering, University of Delaware(电气与计算机工程系,德雷塞尔大学) Swiss Federal Institute for Forest, Snow, and Landscape Research WSL(瑞士森林、雪和景观研究联邦 institute WSL)

专题命中 感知 :LiDAR(abstract);分类 cs.CV、eess.IV

AI总结 MURTreeFormer通过多模态分割框架降低不确定性,提升高分辨率遥感图像中树冠分割的鲁棒性与准确性。

Journal ref IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.10213 2025-12-12 eess.IV cs.RO 62%

Active Optics for Hyperspectral Imaging of Reflective Agricultural Leaf Sensors

主动光学用于反射农业叶片传感器的高光谱成像

Dexter Burns, Sanjeev Koppal

机构 * University of Florida Department of Electrical

专题命中 感知 :LiDAR(abstract);分类 cs.RO、eess.IV

AI总结 本文提出了一种低成本主动光学系统,用于高效定位和采样农业叶片传感器,实现高光谱成像和实时植物健康监测。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08247 2025-12-10 cs.CV cs.AI 62%

Distilling Future Temporal Knowledge with Masked Feature Reconstruction for 3D Object Detection

通过掩码特征重建进行未来时间知识蒸馏以实现3D目标检测

Haowen Zheng, Hu Zhu, Lu Deng, Weihao Gu, Yang Yang, Yanyan Liang

机构 * HAOMO.AI Technology(HAOMO.AI科技公司)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

AI总结 本文提出FTKD方法,通过掩码特征重建和未来引导logit蒸馏,提升3D目标检测中未来帧知识的转移效果,实现性能提升。

Comments AAAI-26

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.17665 2025-12-08 cs.CV cs.RO 62%

Perspective-Invariant 3D Object Detection

视角不变的3D物体检测

Ao Liang, Lingdong Kong, Dongyue Lu, Youquan Liu, Jian Fang, Huaici Zhao, Wei Tsang Ooi

专题命中 感知 :LiDAR(abstract);分类 cs.RO、cs.CV

AI总结 本文提出Pi3DET数据集和跨平台适应框架,实现视角不变的3D物体检测,推动非车辆平台的3D检测研究。

Comments ICCV 2025; 54 pages, 18 figures, 22 tables; Project Page at https://pi3det.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04039 2025-12-04 cs.CV cs.AI cs.LG 62%

Fast & Efficient Normalizing Flows and Applications of Image Generative Models

高效且快速的归一化流及图像生成模型的应用

Sandeep Nagar

机构 * International Institute of Information Technology (Deemed to be University)(国际信息技术学院(认定为大学)) IIIT Hyderabad(IIIT海得拉尔)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

AI总结 本研究提出高效归一化流及图像生成模型的应用,包括超分辨率、农业质量评估、地质制图、隐私保护及艺术修复等

Comments PhD Thesis

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.04096 2025-12-03 cs.RO cs.CV 62%

Image-Based Relocalization and Alignment for Long-Term Monitoring of Dynamic Underwater Environments

基于图像的重定位与对齐用于动态水下环境的长期监测

Beverley Gorry, Tobias Fischer, Michael Milford, Alejandro Fontan

机构 * Queensland University of Technology(昆士兰理工大学)

专题命中 感知 :BEV(abstract);分类 cs.RO、cs.CV

AI总结 本文提出了一种结合VPR、特征匹配和图像分割的方法,用于水下环境的长期监测,并引入了首个大规模水下VPR基准测试。

Journal ref Proceedings of the 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), Hangzhou, China, pp. 10749-10756, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01755 2025-12-02 cs.CV cs.RO 62%

3EED: Ground Everything Everywhere in 3D

3EED: 在三维中万物皆 grounded

Rong Li, Yuhao Dong, Tianshuai Hu, Ao Liang, Youquan Liu, Dongyue Lu, Liang Pan, Lingdong Kong, Junwei Liang, Ziwei Liu

机构 * WorldBench Team(WorldBench团队)

专题命中 感知 :LiDAR(abstract);分类 cs.RO、cs.CV

AI总结 3EED提出一个大规模多平台多模态三维 grounding 基准测试,通过提供丰富的户外场景数据和跨平台学习技术,推动语言驱动的三维具身感知研究。

Comments NeurIPS 2025 DB Track; 38 pages, 17 figures, 10 tables; Project Page at https://project-3eed.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.18016 2025-12-02 cs.RO cs.AI 62%

ADA-DPM: A Neural Descriptors-based Adaptive Noise Filtering Strategy for SLAM

ADA-DPM: 基于神经描述符的自适应噪声过滤策略用于SLAM

Yongxin Shao, Aihong Tan, Binrui Wang, Yinlian Jin, Licong Guan, Peng Liao

机构 * College of Metrology Measurement and Instrument(计量测量与仪器学院) College of Mechanical and Electrical Engineering(机械与电气工程学院)

专题命中 感知 :LiDAR(abstract);分类 cs.RO、cs.AI

AI总结 ADA-DPM通过动态分割、全局重要性评分和跨层图卷积模块,提升SLAM中动态物体干扰和噪声的鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.01791 2025-12-02 cs.CV cs.RO 62%

A Minimal Subset Approach for Informed Keyframe Sampling in Large-Scale SLAM

大规模SLAM中用于有信息关键帧采样的最小子集方法

Nikolaos Stathoulopoulos, Christoforos Kanellakis, George Nikolakopoulos

机构 * Robotics and AI Group, Department of Computer, Electrical and Space Engineering, Luleå University of Technology(机器人与人工智能组,计算机、电气与空间工程系,卢勒奥技术大学)

专题命中 感知 :LiDAR(abstract);分类 cs.RO、cs.CV

AI总结 本文提出了一种基于最小子集方法的在线关键帧采样技术,通过减少冗余和保留信息来提升大规模SLAM中的闭环检测性能和定位精度。

Comments Please cite the published version. 8 pages, 9 figures

Journal ref IEEE Robotics and Automation Letters, vol. 11, no. 1, pp. 738-745, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.23311 2025-12-01 cs.CV cs.AI cs.CL 62%

Toward Automatic Safe Driving Instruction: A Large-Scale Vision Language Model Approach

迈向自动安全驾驶指令:一种大规模视觉语言模型方法

Haruki Sakajo, Hiroshi Takato, Hiroshi Tsutsui, Komei Soda, Hidetaka Kamigaito, Taro Watanabe

机构 * Nara Institute of Science and Technology(奈良科学技术研究所) Teatis inc.(Teatis公司) Queensland university of technology(昆士兰理工大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

AI总结 本文提出了一种大规模视觉语言模型方法,用于生成安全驾驶指令,通过构建数据集并评估模型性能,展示了微调模型在自动驾驶安全中的应用与挑战。

Comments Accepted to MMLoSo 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18668 2025-11-25 cs.CV eess.IV 62%

Data Augmentation Strategies for Robust Lane Marking Detection

用于稳健车道标记检测的数据增强策略

Flora Lian, Dinh Quang Huynh, Hector Penades, J. Stephany Berrio Perez, Mao Shan, Stewart Worrall

机构 * The University of Sydney(悉尼大学) Australian Centre for Robotics(澳大利亚机器人中心)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、eess.IV

AI总结 本文提出基于生成AI的数据增强方法,提升车道标记检测在不同视角下的鲁棒性,通过几何变换、图像修复和车辆叠加模拟部署场景,提升模型在阴影等干扰下的性能。

Comments 8 figures, 2 tables, 10 pages, ACRA, Australasian conference on robotics and automation

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17612 2025-11-25 cs.CV cs.AI 62%

Unified Low-Light Traffic Image Enhancement via Multi-Stage Illumination Recovery and Adaptive Noise Suppression

通过多阶段照明恢复和自适应噪声抑制实现统一的低光照交通图像增强

Siddiqua Namrah

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

AI总结 本文提出了一种无监督的多阶段深度学习框架,通过照明恢复和噪声抑制提升低光照交通图像质量,提高自动驾驶等系统的感知可靠性。

Comments Master's thesis, Korea University, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12638 2025-11-18 cs.CV cs.AI 62%

edgeVLM: Cloud-edge Collaborative Real-time VLM based on Context Transfer

Chen Qian, Xinran Yu, Zewen Huang, Danyang Li, Qiang Ma, Fan Dang, Xuan Ding, Guangyong Shang, Zheng Yang

机构 * Tsinghua University(清华大学) Beijing Jiaotong University(北京交通大学) Inspur Yunzhou Industrial Internet Co., Ltd(Inspur云洲工业互联网有限公司)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11845 2025-11-18 cs.RO cs.AI cs.AR 62%

Autonomous Underwater Cognitive System for Adaptive Navigation: A SLAM-Integrated Cognitive Architecture

K. A. I. N Jayarathne, R. M. N. M. Rathnayaka, D. P. S. S. Peiris

机构 * Department of Computational Mathematics University of Moratuwa(计算数学系大学莫图瓦)

专题命中 感知 :LiDAR(abstract);分类 cs.RO、cs.AI

Comments 6 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11676 2025-11-18 cs.LG cs.AI cs.CV 62%

Learning with Preserving for Continual Multitask Learning

Hanchen David Wang, Siwoo Bae, Zirong Chen, Meiyi Ma

机构 * Vanderbilt University(范德比大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

Comments 25 pages, 16 figures, accepted at AAAI-2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11838 2025-11-18 cs.CV cs.AI 62%

Probabilistic Robustness Analysis in High Dimensional Space: Application to Semantic Segmentation Network

Navid Hashemi, Samuel Sasaki, Diego Manzanas Lopez, Lars Lindemann, Ipek Oguz, Meiyi Ma, Taylor T. Johnson

机构 * Department of Computer Science, Vanderbilt University(计算机科学系,范德比尔特大学) Automatic Control Laboratory, ETH Zürich(自动控制实验室,苏黎世联邦理工学院)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08640 2025-11-13 cs.CV cs.AI 62%

Predict and Resist: Long-Term Accident Anticipation under Sensor Noise

Xingcheng Liu, Bin Rao, Yanchen Guan, Chengyue Wang, Haicheng Liao, Jiaxun Zhang, Chengyu Lin, Meixin Zhu, Zhenning Li

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

Comments accepted by the Fortieth AAAI Conference on Artificial Intelligence (AAAI-26)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07950 2025-11-12 cs.RO cs.AI 62%

USV Obstacles Detection and Tracking in Marine Environments

Yara AlaaEldin, Enrico Simetti, Francesca Odone

机构 * DIBRIS - Department of Computer Science, Bioengineering, Robotics and System Engineering(DIBRIS-计算机科学、生物工程、机器人与系统工程系) University of Genova(热那亚大学)

专题命中 感知 :LiDAR(abstract);分类 cs.RO、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07238 2025-11-11 cs.CV cs.AI 62%

Leveraging Text-Driven Semantic Variation for Robust OOD Segmentation

Seungheon Song, Jaekoo Lee

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

Comments 8 pages, 5 figure references, 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.00527 2025-11-11 eess.IV cs.CV 62%

MAROON: A Dataset for the Joint Characterization of Near-Field High-Resolution Radio-Frequency and Optical Depth Imaging Techniques

Vanessa Wirth, Johanna Bräunig, Nikolai Hofmann, Martin Vossiek, Tim Weyrich, Marc Stamminger

机构 * Friedrich-Alexander-Universität Erlangen-Nürnberg(埃朗根-纽伦堡弗里德里希-亚历山大大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、eess.IV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.07843 2025-11-11 cs.CV cs.RO 62%

Real-time Multi-view Omnidirectional Depth Estimation for Real Scenarios based on Teacher-Student Learning with Unlabeled Data

Ming Li, Xiong Yang, Chaofan Wu, Jiaheng Li, Pinzhi Wang, Xuejiao Hu, Sidan Du, Yang Li

机构 * School of Artificial Intelligence/School of Future Technology, Nanjing University of Information Science and Technology(人工智能学院/未来技术学院,信息科学与技术大学) School of Electronic Science and Engineering, Nanjing University(电子科学与工程学院,南京大学) School of Computer Engineering, Jinling Institute of Technology(计算机工程学院,金陵科技学院) Suzhou High Technology Research Institute, Nanjing University(苏州高新技术研究院,南京大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.RO、cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05404 2025-11-10 cs.CV cs.AI 62%

Multi-modal Loop Closure Detection with Foundation Models in Severely Unstructured Environments

Laura Alejandra Encinar Gonzalez, John Folkesson, Rudolph Triebel, Riccardo Giubilato

专题命中 感知 :LiDAR(abstract);分类 cs.CV、cs.AI

Comments Under review for ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26358 2025-10-31 cs.RO cs.CV 62%

AgriGS-SLAM: Orchard Mapping Across Seasons via Multi-View Gaussian Splatting SLAM

Mirko Usuelli, David Rapado-Rincon, Gert Kootstra, Matteo Matteucci

机构 * Dipartimento di Bioingegneria, Elettronica e Informazione, Politecnico di Milano(生物工程、电子与信息系,米兰理工大学) Agricultural Biosystems Engineering, Wageningen University & Research(农业生物系统工程,瓦赫宁根大学与研究)

专题命中 感知 :LiDAR(abstract);分类 cs.RO、cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22669 2025-10-28 cs.CV cs.AI 62%

LVD-GS: Gaussian Splatting SLAM for Dynamic Scenes via Hierarchical Explicit-Implicit Representation Collaboration Rendering

Wenkai Zhu, Xu Li, Qimin Xu, Benwu Wang, Kun Wei, Yiming Peng, Zihang Wang

机构 * School of Instrument Science and Engineering, Southeast University, Nanjing, China(仪器科学与工程学院,东南大学,南京,中国)

专题命中 感知 :LiDAR(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22243 2025-10-28 cs.CV cs.AI 62%

Real-Time Semantic Segmentation on FPGA for Autonomous Vehicles Using LMIINet with the CGRA4ML Framework

Amir Mohammad Khadem Hosseini, Sattar Mirzakuchaki

机构 * Department of Electrical Engineering, Iran University of Science and Technology (IUST)(电气工程系,伊朗科学技术大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏