arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

自动驾驶

自动驾驶感知、规划、BEV、占用预测、激光雷达和仿真评测。

2025-12-23 至 2025-12-23 共收录 5 信号源:cs.RO, cs.CV, eess.IV, cs.AI

1. 感知 5 篇

2511.13031 2025-12-23 cs.CV 77%

Towards 3D Object-Centric Feature Learning for Semantic Scene Completion

面向语义场景补全的3D物体感知特征学习

Weihua Wang, Yubo Cui, Xiangru Lin, Zhiheng Li, Zheng Fang

专题命中 感知 :autonomous driving(abstract);BEV(abstract);occupancy(abstract);分类 cs.CV

AI总结 本文提出Ocean框架,通过分解场景为独立物体实例,提升语义场景补全的精度和性能。

Comments Accepted to AAAI-2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.20034 2025-12-23 cs.CV cs.RO 62%

NeSLAM: Neural Implicit Mapping and Self-Supervised Feature Tracking With Depth Completion and Denoising

NeSLAM: 基于深度补全与去噪的神经隐式映射与自监督特征跟踪

Tianchen Deng, Yanbo Wang, Hongle Xie, Hesheng Wang, Jingchuan Wang, Danwei Wang, Weidong Chen

机构 * Institute of Medical Robotics and Department of Automation, Shanghai Jiao Tong University(上海交通大学医疗机器人研究所和自动化系) Key Laboratory of System Control and Information Processing, Ministry of Education, Shang hai 200240, China(教育部系统控制与信息处理重点实验室,上海200240,中国) School of Electrical and Electronic Engineering, Nanyang Technological University(南洋理工大学电子与电气工程学院)

专题命中 感知 :occupancy(abstract);分类 cs.RO、cs.CV

AI总结 NeSLAM通过深度补全与去噪网络和SDF表示,实现高精度的3D重建、鲁棒的相机跟踪和视点合成。

Journal ref IEEE Transactions on Automation Science and Engineering, vol. 22, pp. 12309-12321, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18082 2025-12-23 cs.CV cs.AI 62%

Uncertainty-Gated Region-Level Retrieval for Robust Semantic Segmentation

不确定性门控区域级检索用于鲁棒语义分割

Shreshth Rajan, Raymond Liu

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

AI总结 本文提出一种不确定性门控区域级检索机制,提升语义分割在域移位下的鲁棒性和精度,实现分割准确率提升11.3%且检索成本降低87.5%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.15109 2025-12-23 eess.SP cs.AI cs.IT math.IT 57%

Large Model Enabled Embodied Intelligence for 6G Integrated Perception, Communication, and Computation Network

大模型赋能的具身智能用于6G集成感知、通信与计算网络

Zhuoran Li, Zhen Gao, Xinhua Liu, Zheng Wang, Xiaotian Zhou, Lei Liu, Yongpeng Wu, Wei Feng, Yongming Huang

机构 * School of Information and Electronics(信息与电子学院) School of Interdisciplinary Science(交叉科学学院) State Key Laboratory of CNS/ATM(CNS/ATM国家重点实验室) Beijing Institute of Technology(北京理工大学) Advanced Technology Research Institute(先进技术研究院) Yangtze Delta Region Academy(长江三角洲地区学院) School of School of Information Science and Engineering(信息科学与工程学院) Institute of Intelligent Communication Technologies(智能通信技术研究院) Shandong Key Laboratory of Intelligent Communication and Sensing-Computing Integration(智能通信与传感-计算集成山东省重点实验室) Zhejiang Provincial Key Laboratory of Information Processing, Communication and Networking(信息处理、通信与网络浙江省重点实验室) Department of Electronic Engineering(电子工程系) State Key Laboratory of Space Network and Communications(空间网络与通信国家重点实验室) National Mobile Communications Research Laboratory(移动通信研究中心) Purple Mountain Laboratories(紫金山实验室)

专题命中 感知 :autonomous driving(abstract);分类 cs.AI

AI总结 本文提出利用大模型赋能基站实现感知、通信和计算一体化,为6G系统提供安全关键的智能解决方案。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.13982 2025-12-23 cs.CV 57%

FocalComm: Hard Instance-Aware Multi-Agent Perception

FocalComm: 重视困难实例的多智能体感知

Dereje Shenkut, Vijayakumar Bhagavatula

机构 * Carnegie Mellon University(卡内基梅隆大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

AI总结 FocalComm通过聚焦困难实例特征交换,提升多智能体协作感知中对行人等安全关键物体的检测性能。

Comments WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏