arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

自动驾驶

自动驾驶感知、规划、BEV、占用预测、激光雷达和仿真评测。

共收录 6072 信号源:cs.RO, cs.CV, eess.IV, cs.AI

1. 感知 6072 篇

2508.05501 2025-11-19 cs.CV 57%

SMOL-MapSeg: Show Me One Label as prompt

Yunshuang Yuan, Frank Thiemann, Thorsten Dahms, Monika Sester

机构 * IKG

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13552 2025-11-18 cs.CV 57%

TSE-Net: Semi-supervised Monocular Height Estimation from Single Remote Sensing Images

Sining Chen, Xiao Xiang Zhu

机构 * Chair of Data Science in Earth Observation, Technical University of Munich (TUM)(地球观测数据科学教授职位,慕尼黑技术大学) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心)

专题命中 感知 :LiDAR(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13195 2025-11-18 cs.CV 57%

Difficulty-Aware Label-Guided Denoising for Monocular 3D Object Detection

Soyul Lee, Seungmin Baek, Dongbo Min

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

Comments AAAI 2026 accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13138 2025-11-18 cs.CV 57%

WinMamba: Multi-Scale Shifted Windows in State Space Model for 3D Object Detection

Longhui Zheng, Qiming Xia, Xiaolu Chen, Zhaoliang Liu, Chenglu Wen

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

Comments 9 pages, 3 figures,

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13055 2025-11-18 cs.CV 57%

Monocular 3D Lane Detection via Structure Uncertainty-Aware Network with Curve-Point Queries

Ruixin Liu, Zejian Yuan

专题命中 感知 :BEV(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12778 2025-11-18 cs.RO 57%

DR. Nav: Semantic-Geometric Representations for Proactive Dead-End Recovery and Navigation

Vignesh Rajagopal, Kasun Weerakoon Kulathun Mudiyanselage, Gershom Devake Seneviratne, Pon Aswin Sankaralingam, Mohamed Elnoor, Jing Liang, Rohan Chandra, Dinesh Manocha

机构 * Dept. of Computer Science at the University of Virginia(弗吉尼亚大学计算机科学系) Dept. of Computer Science at the University of Maryland College Park(马里兰大学学院市计算机科学系)

专题命中 感知 :LiDAR(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12648 2025-11-18 cs.CR cs.AI cs.LG 57%

Scalable Hierarchical AI-Blockchain Framework for Real-Time Anomaly Detection in Large-Scale Autonomous Vehicle Networks

Rathin Chandra Shit, Sharmila Subudhi

机构 * organization= Dept. of Computer Science \& Engg., International Institute of Information Technology , city= Bhubaneswar , postcode= 751003 , state= Odisha , country= India organization= Dept. of Computer Science, Maharaja Sriram Chandra Bhanja Deo University , city= Baripada , postcode= 757003 , state= Odisha , country= India

专题命中 感知 :autonomous driving(abstract);分类 cs.AI

Comments Submitted to the Journal

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00489 2025-11-18 cs.CV 57%

Density-aware global-local attention network for point cloud segmentation

Chade Li, Pengju Zhang, Jiaming Zhang, Yihong Wu

机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences, Beijing 100190, China(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院,北京100190,中国) School of Artificial Intelligence, University of Chinese Academy of Sciences, Beijing 100049, China(人工智能学院,中国科学院大学,北京100049,中国)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

Comments Accepted by Image and Vision Computing

Journal ref Image and Vision Computing, Volume 165, 2026, 105822

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.00620 2025-11-18 cs.CV 57%

Lane Graph Extraction from Aerial Imagery via Lane Segmentation Refinement with Diffusion Models

Antonio Ruiz, Andrew Melnik, Nicolo Savioli, Dong Wang, Yanfeng Zhang, Helge Ritter

机构 * Riemann Lab , Huawei(里曼实验室,华为) Center for Cognitive Interaction Technology (CITEC), Faculty of Technology, Bielefeld University(认知交互技术中心(CITEC),技术学院,比勒菲尔德大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

Journal ref Remote Sensing, 17(16), 2845

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09955 2025-11-14 cs.CV 57%

Robust Object Detection with Pseudo Labels from VLMs using Per-Object Co-teaching

Uday Bhaskar, Rishabh Bhattacharya, Avinash Patel, Sarthak Khoche, Praveen Anil Kulkarni, Naresh Manwani

机构 * Machine Learning Lab IIIT Hyderabad(IIIT Hyderabad 机器学习实验室) Bosch Global Software Technologies(博世全球软件技术公司)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08741 2025-11-14 cs.RO cs.SY eess.SY 57%

ATOM-CBF: Adaptive Safe Perception-Based Control under Out-of-Distribution Measurements

Kai S. Yun, Navid Azizan

机构 * Massachusetts Institute of Technology(麻省理工学院)

专题命中 感知 :LiDAR(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.04326 2025-11-13 cs.CV 57%

LMSeg: An end-to-end geometric message-passing network on barycentric dual graphs for large-scale landscape mesh segmentation

Zexian Huang, Kourosh Khoshelham, Martin Tomko

机构 * The University of Melbourne(墨尔本大学)

专题命中 感知 :LiDAR(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07966 2025-11-12 cs.CV 57%

Multi-Modal Assistance for Unsupervised Domain Adaptation on Point Cloud 3D Object Detection

Shenao Zhao, Pengpeng Liang, Zhoufan Yang

专题命中 感知 :LiDAR(abstract);分类 cs.CV

Comments Accepted to AAAI-26

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07862 2025-11-12 cs.CV 57%

MonoCLUE : Object-Aware Clustering Enhances Monocular 3D Object Detection

Sunghun Yang, Minhyeok Lee, Jungho Lee, Sangyoun Lee

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

Comments Accepted to AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11171 2025-11-12 cs.CV 57%

SPHERE: Semantic-PHysical Engaged REpresentation for 3D Semantic Scene Completion

Zhiwen Yang, Yuxin Peng

机构 * Peking University(北京大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

Comments 10 pages, 6 figures, accepted by ACM MM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06632 2025-11-11 cs.CV 57%

DIAL-GS: Dynamic Instance Aware Reconstruction for Label-free Street Scenes with 4D Gaussian Splatting

Chenpeng Su, Wenhua Wu, Chensheng Peng, Tianchen Deng, Zhe Liu, Hesheng Wang

机构 * Chenpeng Su ∗ , Wenhua Wu ∗ , Chensheng Peng, Tianchen Deng, Zhe Liu, Hesheng Wang †(无明确机构)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.10413 2025-11-11 cs.CV 57%

HyCTAS: Multi-Objective Hybrid Convolution-Transformer Architecture Search for Real-Time Image Segmentation

Hongyuan Yu, Cheng Wan, Xiyang Dai, Mengchen Liu, Dongdong Chen, Bin Xiao, Yan Huang, Yuan Lu, Liang Wang

机构 * The Multimedia Department, Xiaomi Inc. (XIAOMI)(小米公司多媒体部门) Georgia Institute of Technology (GATECH)(佐治亚理工学院) Center for Research on Intelligent Perception and Computing (CRIPAC)(智能感知与计算研究中心) National Laboratory of Pattern Recognition (NLPR)(模式识别国家实验室) University of Chinese Academy of Sciences (UCAS)(中国科学院大学) Microsoft Inc.(微软公司)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

Comments 24 pages, 5 figures, published at Neurocomputing

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06408 2025-11-11 cs.CV 57%

VDNeRF: Vision-only Dynamic Neural Radiance Field for Urban Scenes

Zhengyu Zou, Jingfeng Li, Hao Li, Xiaolei Hou, Jinwen Hu, Jingkun Chen, Lechao Cheng, Dingwen Zhang

机构 * Northwestern Polytechnical University(西北工业大学) University of Oxford(牛津大学) Hefei University Of Technology(合肥工业大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06337 2025-11-11 cs.CV 57%

BuildingWorld: A Structured 3D Building Dataset for Urban Foundation Models

Shangfeng Huang, Ruisheng Wang, Xin Wang

专题命中 感知 :LiDAR(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05108 2025-11-10 cs.CV 57%

SnowyLane: Robust Lane Detection on Snow-covered Rural Roads Using Infrastructural Elements

Jörg Gamerdinger, Benedict Wetzel, Patrick Schulz, Sven Teufel, Oliver Bringmann

机构 * University of Tübingen, Faculty of Science, Department of Computer Science, Embedded Systems Group(图宾根大学科学学院计算机科学系嵌入式系统组) Forschungszentrum Informatik (FZI) Karlsruhe(信息研究所(FZI)卡尔斯鲁厄)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.04220 2025-11-06 cs.CV 57%

Struct2D: A Perception-Guided Framework for Spatial Reasoning in MLLMs

Fangrui Zhu, Hanhui Wang, Yiming Xie, Jing Gu, Tianye Ding, Jianwei Yang, Huaizu Jiang

机构 * Northeastern University(东北大学) Microsoft Research(微软研究院) University of Southern California(南加州大学) University of California, Santa Cruz(加州大学圣克鲁兹分校)

专题命中 感知 :BEV(abstract);分类 cs.CV

Comments NeurIPS 2025, code link: https://github.com/neu-vi/struct2d

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.09915 2025-11-05 cs.CV 57%

Crucial-Diff: A Unified Diffusion Model for Crucial Image and Annotation Synthesis in Data-scarce Scenarios

Siyue Yao, Mingjie Sun, Eng Gee Lim, Ran Yi, Baojiang Zhong, Moncef Gabbouj

机构 * School of Advanced Technology, Xi’an Jiaotong-Liverpool University(西安交通大学利物浦大学先进科技学院) School of Computer Science and Technology, Soochow University(苏州大学计算机科学与技术学院) School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机科学学院) Faculty of Information Technology and Communication Sciences, Tampere University(塔尔库大学信息科技与通信科学学院)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

Comments Accepted by IEEE Transactions on Image Processing (TIP), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07215 2025-11-05 cs.RO cs.MM 57%

RoboTron-Mani: All-in-One Multimodal Large Model for Robotic Manipulation

Feng Yan, Fanfan Liu, Liming Zheng, Yufeng Zhong, Yiyang Huang, Zechao Guan, Chengjian Feng, Lin Ma

机构 * Meituan(美团)

专题命中 感知 :occupancy(abstract);分类 cs.RO

Journal ref Proceedings of the IEEE/CVF International Conference on Computer Vision 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.13916 2025-11-04 cs.RO 57%

Robotic Monitoring of Colorimetric Leaf Sensors for Precision Agriculture

Malakhi Hopkins, Alice Kate Li, Shobhita Kramadhati, Jackson Arnold, Akhila Mallavarapu, Chavez F. K. Lawrence, Anish Bhattacharya, Varun Murali, Sanjeev J. Koppal, Cherie R. Kagan, Vijay Kumar

机构 * GRASP Laboratory, University of Pennsylvania(宾夕法尼亚大学GRASP实验室) University of Florida(佛罗里达大学) Amazon Robotics(亚马逊机器人)

专题命中 感知 :LiDAR(abstract);分类 cs.RO

Comments Revised version. Initial version was accepted to the Novel Approaches for Precision Agriculture and Forestry with Autonomous Robots IEEE ICRA Workshop - 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00653 2025-11-04 cs.CV 57%

Benchmarking individual tree segmentation using multispectral airborne laser scanning data: the FGI-EMIT dataset

Lassi Ruoppa, Tarmo Hietala, Verneri Seppänen, Josef Taher, Teemu Hakala, Xiaowei Yu, Antero Kukko, Harri Kaartinen, Juha Hyyppä

机构 * Finnish Geospatial Research Institute FGI(芬兰地理空间研究 institute) The National Land Survey of Finland(芬兰国家土地调查局) Department of Remote Sensing and Photogrammetry(遥感与摄影测量系) Department of Built Environment, School of Engineering, Aalto University(环境学院,工程学院,阿alto大学)

专题命中 感知 :LiDAR(abstract);分类 cs.CV

Comments 39 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00073 2025-11-04 cs.CV 57%

Habitat and Land Cover Change Detection in Alpine Protected Areas: A Comparison of AI Architectures

Harald Kristen, Daniel Kulmer, Manuela Hirschmugl

机构 * University of Graz(格拉茨大学) Joanneum Research(乔安姆研究)

专题命中 感知 :LiDAR(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00026 2025-11-04 cs.RO 57%

Gen AI in Automotive: Applications, Challenges, and Opportunities with a Case study on In-Vehicle Experience

Chaitanya Shinde, Divya Garikapati

专题命中 感知 :autonomous driving(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26833 2025-11-03 cs.CR cs.AI cs.LG 57%

VISAT: Benchmarking Adversarial and Distribution Shift Robustness in Traffic Sign Recognition with Visual Attributes

Simon Yu, Peilin Yu, Hongbo Zheng, Huajie Shao, Han Zhao, Lui Sha

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Brown University(布朗大学) William & Mary(威廉与玛丽学院)

专题命中 感知 :autonomous driving(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.25146 2025-10-30 cs.CV 57%

EA3D: Online Open-World 3D Object Extraction from Streaming Videos

Xiaoyu Zhou, Jingqi Wang, Yuang Jia, Yongtao Wang, Deqing Sun, Ming-Hsuan Yang

机构 * Wangxuan Institute of Computer Technology, Peking University(北京大学王轩计算机技术研究所) Google DeepMind(谷歌DeepMind) University of California, Merced(加州大学梅尔塞德斯分校)

专题命中 感知 :occupancy(abstract);分类 cs.CV

Comments The Thirty-Ninth Annual Conference on Neural Information Processing Systems(NeurIPS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12015 2025-10-30 cs.CV 57%

InstDrive: Instance-Aware 3D Gaussian Splatting for Driving Scenes

Hongyuan Liu, Haochen Yu, Bochao Zou, Jianfei Jiang, Qiankun Liu, Jiansheng Chen, Huimin Ma

机构 * University of Science and Technology Beijing(北京科技大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏