arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

2025-12-09 至 2025-12-09 共收录 10 信号源:cs.CV, cs.GR, cs.RO

1. 点云 10 篇

2511.11210 2025-12-09 cs.CV 83%

STONE: Pioneering the One-to-N Universal Backdoor Threat in 3D Point Cloud

STONE:开创性地探索3D点云中的一对多通用后门威胁

Dongmei Shan, Wei Lian, Chongxia Wang

机构 * Changzhi University(长治大学)

专题命中 点云 :point cloud(title,abstract);3D vision(abstract);分类 cs.CV

AI总结 STONE首次提出3D点云中一对多通用后门威胁,通过可配置球形触发器设计实现高攻击成功率,为多目标后门威胁提供理论与实证基础。

Comments 15 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07040 2025-12-09 cs.CV cs.CR 79%

3D-ANC: Adaptive Neural Collapse for Robust 3D Point Cloud Recognition

3D-ANC:适应性神经坍塌用于鲁棒的3D点云识别

Yuanmin Huang, Wenxuan Li, Mi Zhang, Xiaohan Zhang, Xiaoyu You, Min Yang

专题命中 点云 :point cloud(title,abstract);分类 cs.CV

AI总结 3D-ANC通过神经坍塌机制解决3D点云识别中的对抗攻击问题,结合ETF对齐和自适应训练框架提升模型鲁棒性。

Comments AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06882 2025-12-09 cs.CV 79%

Hierarchical Image-Guided 3D Point Cloud Segmentation in Industrial Scenes via Multi-View Bayesian Fusion

基于多视角贝叶斯融合的工业场景分层图像引导3D点云分割

Yu Zhu, Naoya Chiba, Koichi Hashimoto

专题命中 点云 :point cloud(title,abstract);分类 cs.CV

AI总结 本文提出一种分层图像引导的3D点云分割框架,通过多视角贝叶斯融合解决工业场景中的遮挡和语义不一致问题,提升分割精度与鲁棒性。

Comments Accepted to BMVC 2025 (Sheffield, UK, Nov 24-27, 2025). Supplementary video and poster available upon request

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06058 2025-12-09 cs.CV 79%

Representation Learning for Point Cloud Understanding

点云理解的表示学习

Siming Yan

机构 * The University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 点云 :point cloud(title,abstract);分类 cs.CV

AI总结 本研究提出整合预训练2D模型以提升3D点云理解的方法,通过自监督学习和迁移学习改进点云表示学习。

Comments 181 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07211 2025-12-09 cs.CV 74%

Object Pose Distribution Estimation for Determining Revolution and Reflection Uncertainty in Point Clouds

点云中确定旋转和反射不确定性时的物体姿态分布估计

Frederik Hagelskjær, Dimitrios Arapis, Steffen Madsen, Thorbjørn Mosekjær Iversen

机构 * SDU Robotics(SDU机器人研究所) The Mærsk Mc-Kinney Møller Institute(马士基麦金尼莫勒研究所) University of Southern Denmark(南部丹麦大学)

专题命中 点云 :point cloud(title);分类 cs.CV

AI总结 本文提出了一种基于深度学习的点云姿态分布估计方法,无需RGB输入,用于评估旋转和反射的不确定性。

Comments 8 pages, 8 figures, 5 tables, ICCR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19912 2025-12-09 cs.CV cs.LG cs.RO 62%

Enhanced Spatiotemporal Consistency for Image-to-LiDAR Data Pretraining

增强的时空一致性用于图像到LiDAR数据预训练

Xiang Xu, Lingdong Kong, Hui Shuai, Wenwei Zhang, Liang Pan, Kai Chen, Ziwei Liu, Qingshan Liu

机构 * College of Computer Science and Technology, Nanjing University of Aeronautics and Astronautics(南京航空航天大学计算机科学与技术学院) School of Computing, Department of Computer Science, National University of Singapore(新加坡国立大学计算机学院) School of Computer Science, Nanjing University of Posts and Telecommunications(南京邮电大学计算机学院) Shanghai AI Laboratory(上海人工智能实验室) S-Lab, Nanyang Technological University(南洋理工大学S实验室)

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

AI总结 SuperFlow++通过整合时空线索提升图像到LiDAR数据预训练效果,实现更鲁棒的特征表示和更高效的自动驾驶感知。

Comments IEEE Transactions on Pattern Analysis and Machine Intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07599 2025-12-09 cs.CV 57%

Online Segment Any 3D Thing as Instance Tracking

在线分割三维物体作为实例跟踪

Hanshi Wang, Zijian Cai, Jin Gao, Yiwei Zhang, Weiming Hu, Ke Wang, Zhipeng Zhang

机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), CASIA(多模态人工智能系统国家重点实验室(MAIS),中国科学院自动化所) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) AutoLab, School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院AutoLab) Anyverse Intelligence Beijing Key Laboratory of Super Intelligent Security of Multi-Modal Information(北京超智能多模态信息安全重点实验室) School of Information Science and Technology, ShanghaiTech University(上海科技大学信息科学与技术学院)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 将在线3D分割重新定义为实例跟踪问题,通过时间信息传播和空间一致性学习提升具身智能体对环境的理解能力。

Comments NeurIPS 2025, Code is at https://github.com/AutoLab-SAI-SJTU/AutoSeg3D

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.09620 2025-12-09 cs.CV cs.AI cs.CL 57%

Exploring the Potential of Encoder-free Architectures in 3D LMMs

探索无编码器架构在3D大语言模型中的潜力

Yiwen Tang, Zoey Guo, Zhuhao Wang, Ray Zhang, Qizhi Chen, Junli Liu, Delin Qu, Zhigang Wang, Dong Wang, Bin Zhao, Xuelong Li

机构 * Northwestern Polytechnical University(西北工业大学) Shanghai AI Laboratory(上海人工智能实验室) The Chinese University of Hong Kong(香港中文大学) Tsinghua University(清华大学) Tele AI

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本文提出首个无编码器3D大语言模型ENEL,通过嵌入语义编码和分层几何聚合策略,在3D理解任务中达到与SOTA模型相当的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06517 2025-12-09 cs.RO 57%

Vision-Guided Grasp Planning for Prosthetic Hands in Unstructured Environments

面向无结构环境的假手视觉引导抓取规划

Shifa Sulaiman, Akash Bachhar, Ming Shen, Simon Bøgh

机构 * Control and Automation Section, Department of Electronics Systems, Aalborg University, Denmark(奥尔堡大学电子系统系自动化与自动化部门) Department of Mechanical Engineering, National Institute of Technology,Durgapur, India(印度德里加尔帕国家理工学院机械工程系) The Technical Faculty of IT(信息技术技术学院) Millimeter-Wave Systems, Department of Electronics Systems, Aalborg University, Denmark(毫米波系统,电子系统系,奥尔堡大学,丹麦)

专题命中 点云 :point cloud(abstract);分类 cs.RO

AI总结 本文提出一种基于视觉引导的假手抓取算法,通过整合感知、规划和控制,实现灵活的抓取操作,并在仿真和实验中验证其在无结构环境中的适应性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.09274 2025-12-09 cs.CV 57%

FLARES: Fast and Accurate LiDAR Multi-Range Semantic Segmentation

FLARES: 快速且准确的LiDAR多范围语义分割

Bin Yang, Alexandru Paul Condurache

机构 * Automated Driving Research, Robert Bosch GmbH(罗伯特博世集团自动化驾驶研究部) Institute for Signal Processing, University of Lübeck(吕贝克大学信号处理研究所)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 FLARES通过多范围图像训练和定制的数据增强技术,提升LiDAR语义分割的准确性和效率,实现mIoU提升和推理速度提升。

Comments The paper was accepted by WACV2026

详情

展开后加载摘要…

URL PDF HTML 收藏