arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

自动驾驶

自动驾驶感知、规划、BEV、占用预测、激光雷达和仿真评测。

2026-01-14 至 2026-01-14 共收录 13 信号源:cs.RO, cs.CV, eess.IV, cs.AI

1. 感知 5 篇

2411.15016 2026-01-14 cs.CV cs.RO 84%

MSSF: A 4D Radar and Camera Fusion Framework With Multi-Stage Sampling for 3D Object Detection in Autonomous Driving

MSSF: 一种基于4D雷达和摄像头的多阶段采样融合框架用于自动驾驶中的3D目标检测

Hongsi Liu, Jun Liu, Guangfeng Jiang, Xin Jin

机构 * Department of Electronic Engineering and Information Science, University of Science and Technology of China(电子工程与信息科学系,中国科学技术大学) Ningbo Institute of Digital Twin, Eastern Institute of Technology(宁波数字孪生研究所,东部技术研究所)

专题命中 感知 :autonomous driving(title,abstract);LiDAR(abstract);分类 cs.RO、cs.CV

AI总结 MSSF通过多阶段采样融合4D雷达和摄像头数据,提升自动驾驶中3D目标检测的精度与鲁棒性。

Comments T-TITS accepted, code avaliable

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08042 2026-01-14 physics.app-ph cs.RO 57%

μDopplerTag: CNN-Based Drone Recognition via Cooperative Micro-Doppler Tagging

μDopplerTag: 基于协作微多普勒标记的CNN无人机识别

O. Yerushalimov, D. Vovchuk, A. Glam, P. Ginzburg

专题命中 感知 :LiDAR(abstract);分类 cs.RO

AI总结 本文提出基于电磁标签和CNN的无人机识别方法,利用微多普勒签名实现远距离高精度分类,适用于空域监控等关键应用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12735 2026-01-14 cs.CV 57%

Backdoor Attacks on Open Vocabulary Object Detectors via Multi-Modal Prompt Tuning

通过多模态提示调优对开放词汇目标检测器进行后门攻击

Ankita Raj, Chetan Arora

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

AI总结 TrAP通过多模态提示调优对开放词汇目标检测器实施后门攻击,利用轻量级提示标记植入恶意行为,提升攻击成功率并改进下游任务性能。

Comments Accepted to AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07168 2026-01-14 cs.CV 57%

HisTrackMap: Global Vectorized High-Definition Map Construction via History Map Tracking

HisTrackMap: 通过历史地图追踪构建全局向量高精度地图

Jing Yang, Sen Yang, Xiao Tan, Hanli Wang

机构 * Tongji University(同济大学) Baidu Inc.(百度公司)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

AI总结 HisTrackMap通过历史地图追踪构建全局向量高精度地图,提升时间连续性和几何构造质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.18124 2026-01-14 cs.LG eess.SP stat.ML 50%

Bayesian Multiobject Tracking With Neural-Enhanced Motion and Measurement Models

基于神经网络的贝叶斯多目标跟踪

Shaoxiu Wei, Mingchao Liang, Florian Meyer

机构 * Department of Electrical and Computer Engineering, University of California San Diego(电气与计算机工程系,加州大学圣地亚哥分校) Scripps Institution of Oceanography(斯克里普斯海洋研究所)

专题命中 感知 :autonomous driving(abstract)

AI总结 本文提出一种结合神经网络与贝叶斯方法的多目标跟踪框架,通过增强统计模型提升预测和更新性能,实现在自动驾驶数据集上的最佳表现。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 规划控制 2 篇

2505.20665 2026-01-14 cs.CV 79%

DriveRX: A Vision-Language Reasoning Model for Cross-Task Autonomous Driving

DriveRX:一种用于跨任务自动驾驶的视觉-语言推理模型

Muxi Diao, Lele Yang, Hongbo Yin, Zhexu Wang, Yejie Wang, Daxin Tian, Kongming Liang, Zhanyu Ma

机构 * Beijing University of Posts and Telecommunications(北京邮电大学) Zhongguancun Academy(中关村学院) Beihang University(北航)

专题命中 规划控制 :autonomous driving(title,abstract);分类 cs.CV

AI总结 DriveRX是一种用于自动驾驶的视觉-语言推理模型,通过结构化推理提升多阶段决策能力,并在复杂驾驶条件下表现优异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08259 2026-01-14 cs.NI 50%

Unleashing Tool Engineering and Intelligence for Agentic AI in Next-Generation Communication Networks

释放工具工程与智能以推动下一代通信网络中的智能体AI

Yinqiu Liu, Ruichen Zhang, Dusit Niyato, Abbas Jamalipour, Trung Q. Duong, Dong In Kim

专题命中 规划控制 :trajectory planning(abstract)

AI总结 本文提出通过工具工程和智能提升下一代通信网络中的智能体AI,展示工具智能在无人飞行器轨迹规划中的应用,并提供6G时代的智能体构建路线图。

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 端到端驾驶 1 篇

2601.07966 2026-01-14 cs.LG cond-mat.mtrl-sci 50%

DataScribe: An AI-Native, Policy-Aligned Web Platform for Multi-Objective Materials Design and Discovery

DataScribe:面向多目标材料设计与发现的AI原生、政策对齐的Web平台

Divyanshu Singh, Doguhan Sarıtürk, Cameron Lea, Md Shafiqul Islam, Raymundo Arroyave, Vahid Attari

机构 * organization= Department of Computer Science, Texas A\&M University , city= College Station , state= TX , postcode= 77843 , country= USA organization= Department of Materials Science Engineering, Texas A\&M University , city= College Station , state= TX , postcode= 77843 , country= USA

专题命中 端到端驾驶 :self-driving(abstract)

AI总结 DataScribe是一个AI原生、政策对齐的Web平台,通过统一异构数据、整合多目标优化和可解释性,实现材料设计与发现的闭环流程。

详情

展开后加载摘要…

URL PDF HTML 收藏

4. BEV与占用 1 篇

2408.15235 2026-01-14 cs.CV 57%

Learning-based Multi-View Stereo: A Survey

基于学习的多视图立体:综述

Fangjinhua Wang, Qingtian Zhu, Di Chang, Quankai Gao, Junlin Han, Tong Zhang, Richard Hartley, Marc Pollefeys

机构 * Department of Computer Science, ETH Zurich(苏黎世联邦理工学院计算机科学系) Graduate School of Information Science and Technology, The University of Tokyo(东京大学信息科学与技术研究生院) Department of Computer Science, University of Southern California(南加州大学计算机科学系) Department of Engineering Science, University of Oxford(牛津大学工程科学系) University of Chinese Academy of Sciences(中国科学院大学) School of Computer and Communication Sciences, EPFL(苏黎世联邦理工学院计算机与通信科学学院) Australian National University(澳大利亚国立大学) Microsoft, Zurich(微软(瑞士))

专题命中 BEV与占用 :autonomous driving(abstract);分类 cs.CV

AI总结 本文综述了基于学习的多视图立体方法,重点介绍了基于深度图的方法,并讨论了该领域未来的研究方向。

Comments Accepted to IEEE T-PAMI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 激光雷达 2 篇

2601.01210 2026-01-14 cs.CV cs.RO 81%

Real-Time LiDAR Point Cloud Densification for Low-Latency Spatial Data Transmission

实时LiDAR点云密集化用于低延迟空间数据传输

Kazuhiko Murasaki, Shunsuke Konagai, Masakatsu Aoki, Taiga Yoshida, Ryuichi Tanida

机构 * NTT Human Informatics Laboratories(NTT人机信息实验室)

专题命中 激光雷达 :LiDAR(title,abstract);分类 cs.RO、cs.CV

AI总结 本文提出了一种基于卷积神经网络的实时LiDAR点云密集化方法,实现低延迟的3D场景生成与深度补全,效率比传统方法快15倍。

Journal ref 19th International Conference on Machine Vision Applications (MVA2025), IEICE Transactions on Information and Systems letter

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08420 2026-01-14 cs.CV 57%

MMLGNet: Cross-Modal Alignment of Remote Sensing Data using CLIP

MMLGNet: 利用CLIP实现遥感数据的跨模态对齐

Aditya Chaudhary, Sneha Barman, Mainak Singha, Ankit Jha, Girish Mishra, Biplab Banerjee

专题命中 激光雷达 :LiDAR(abstract);分类 cs.CV

AI总结 MMLGNet通过CLIP实现遥感数据的跨模态对齐,利用多模态语言引导网络有效融合光谱、空间和几何信息,提升语义理解能力。

Comments Accepted at InGARSS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

6. 仿真评测 1 篇

2506.05280 2026-01-14 cs.CV 57%

Unifying Appearance Codes and Bilateral Grids for Driving Scene Gaussian Splatting

统一外观代码与双侧网格用于驾驶场景高斯点漂浮

Nan Wang, Yuantao Chen, Lixing Xiao, Weiqing Xiao, Bohan Li, Zhaoxi Chen, Chongjie Ye, Shaocong Xu, Saining Zhang, Ziyang Yan, Pierre Merriaux, Lei Lei, Tianfan Xue, Hao Zhao

机构 * BAAI(北京人工智能研究院) AIR, THU(清华大学人工智能研究院) SJTU(上海交通大学) EIT(Ningbo)(宁波工程学院) CUHK(香港大学) LeddarTech

专题命中 仿真评测 :autonomous driving(abstract);分类 cs.CV

AI总结 本文提出了一种多尺度双侧网格方法,统一了外观代码与双侧网格,提升了自动驾驶场景中的几何重建精度。

Comments Accepted to NeurIPS 2025 ; Project page: https://bigcileng.github.io/bilateral-driving ; Code: https://github.com/BigCiLeng/bilateral-driving

详情

展开后加载摘要…

URL PDF HTML 收藏

7. 其他自动驾驶 1 篇

2601.08185 2026-01-14 cond-mat.mtrl-sci cs.AI cs.LG cs.MA physics.comp-ph 57%

Autonomous Materials Exploration by Integrating Automated Phase Identification and AI-Assisted Human Reasoning

通过整合自动化相识别和AI辅助的人类推理实现自主材料探索

Ming-Chiang Chang, Maximilian Amsler, Duncan R. Sutherland, Sebastian Ament, Katie R. Gann, Lan Zhou, Louisa M. Smieska, Arthur R. Woll, John M. Gregoire, Carla P. Gomes, R. Bruce van Dover, Michael O. Thompson

机构 * Department of Materials Science and Engineering, Cornell University, Ithaca, NY 14853, United States(材料科学与工程系,康奈尔大学,Ithaca, NY 14853, United States) Department of Computer Science, Cornell University, Ithaca, NY 14853, United States(计算机科学系,康奈尔大学,Ithaca, NY 14853, United States) Joint Center for Artificial Photosynthesis, California Institute of Technology, Pasadena, CA 91125(人工光合作研究中心,加州理工学院,Pasadena, CA 91125) Cornell High Energy Synchrotron Source, Cornell University, Ithaca, NY 14850, United States(康奈尔高能同步辐射源,康奈尔大学,Ithaca, NY 14850, United States)

专题命中 其他自动驾驶 :self-driving(abstract);分类 cs.AI

AI总结 通过整合自动化相识别和AI辅助的人类推理,实现自主材料探索,提升材料合成效率和发现新材料的能力。

Comments Main manuscript: 21 pages(including references), 6 figures. Supplementary Information: 12 pages, 9 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏