arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

2026-08-25 至 2026-08-25 共收录 12 信号源:cs.CV, cs.GR, cs.RO

1. 三维重建 1 篇

2601.14492 2026-08-25 cs.RO 版本更新 57%

UNCLE-Grasp: A Task-Adapted Framework for Uncertainty-Aware Grasping of Leaf-Occluded Strawberries

UNCLE-Grasp: 有不确定性意识的叶片遮挡草莓抓取

Malak Mansour, Ali Abouzeid, Zezhou Sun, Qinbo Sun, Dezhen Song, Abdalla Swikir

机构 * Department of Robotics, Mohamed bin Zayed University of Artificial Intelligence(机器人系,Mohamed bin Zayed人工智能大学)

专题命中 三维重建 :point cloud(abstract);分类 cs.RO

AI总结 UNCLE-Grasp通过建模遮挡和形状补全的几何不确定性,提出了一种在部分遮挡下可靠抓取草莓的方法,通过多假设评估和保守置信界标准提升抓取可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. NeRF 1 篇

2605.15324 2026-08-25 eess.SP 版本更新 80%

P-WRFGS: Pruning 3D Gaussians for Efficient Wireless Radiance Field Construction

Eff-WRFGS:基于3D高斯点划的高效无线辐射场

Chenghong Bian, Meng Hua, Deniz Gunduz

专题命中 NeRF :NeRF(abstract,abstract_cn);Gaussian Splatting(abstract);point cloud(abstract)

AI总结 本文提出Eff-WRFGS框架,利用3D高斯点划建模无线信道,通过可学习掩码优化高斯体重要性,实现存储减少和渲染加速。

Comments Accepted to IEEE Wireless Communication Letter

详情

展开后加载摘要…

URL PDF HTML 收藏

3. Gaussian Splatting 3 篇

2508.03077 2026-08-25 cs.CV 版本更新 90%

RobustGS: Unified Boosting of Feedforward 3D Gaussian Splatting under Low-Quality Conditions

RobustGS:低质量条件下前馈三维高斯溅射的统一增强方法

Anran Wu, Long Peng, Xin Di, Xueyuan Dai, Chen Wu, Yang Wang, Xueyang Fu, Yang Cao, Zheng-Jun Zha

专题命中 Gaussian Splatting :3DGS(summary_cn,abstract);Gaussian Splatting(title,abstract);3D reconstruction(abstract);分类 cs.CV

AI总结 本文提出RobustGS模块,即插即用集成到现有前馈3DGS管道,通过广义退化学习器和语义感知状态空间模型,提升低质量条件下三维重建的鲁棒性与质量,实验达最优效果。

Comments This paper is being withdrawn as the content has been determined, upon review, to be subject to the confidentiality regulations of the relevant institution and must be handled accordingly

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22225 2026-08-25 cs.CV cs.AI 版本更新 85%

ExtrinSplat: Decoupling Geometry and Semantics for Open-Vocabulary Understanding in 3D Gaussian Splatting

ExtrinSplat:解耦几何与语义以实现3D高斯散射中的开放词汇理解

Jiayu Ding, Xinpeng Liu, Zhiyi Pan, Shiqiang Long, Ge Li

机构 * Guangdong Provincial Key Laboratory of Ultra High Definition Immersive Media Technology, Shenzhen Graduate School, Peking University(广东省超高清沉浸式媒体技术重点实验室,北京大学深圳研究生院) School of Computer Science and Technology, Tianjin University(天津大学计算机科学与技术学院) Guangdong Bohua UHD Innovation Center Co., Ltd.(广东博华超高清创新中心有限公司)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract,abstract_cn);分类 cs.CV

AI总结 ExtrinSplat通过解耦几何与语义,利用视觉语言模型生成轻量文本假设,提升3D高斯散射中开放词汇物体选择和语义分割的性能和效率。

Comments Accepted to CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.01844 2026-08-25 cs.CV 版本更新 79%

FaCT-GS: Fast and Scalable CT Reconstruction with Gaussian Splatting

FaCT-GS:基于高斯点撒的快速可扩展CT重建

Pawel Tomasz Pieta, Rasmus Juul Pedersen, Sina Borgi, Jakob Sauer Jørgensen, Jens Wenzel Andreasen, Vedrana Andersen Dahl

机构 * Technical University of Denmark(丹麦技术大学)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);分类 cs.CV

AI总结 本文提出FaCT-GS框架,通过优化体素化和光栅化流程,实现快速且可扩展的CT重建,比现有方法快4-13倍,适用于不同投影尺寸。

Comments Accepted at the European Conference on Computer Vision (ECCV) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 点云 2 篇

2608.19232 2026-08-25 cs.NE cs.AI cs.LG 版本更新 78%

Active Spiking Perception: The Membrane Potential as a Belief State for Anytime 3D Point Cloud Recognition

主动脉冲感知:作为任何时刻3D点云识别置信状态的膜电位

Akarsh Jain, Arya Pawa, Ayush Debnath, Smera Rawal, Sayeed Shafayet Chowdhury

专题命中 点云 :point cloud(title,abstract)

AI总结 提出主动脉冲感知(ASP),以LIF膜电位为置信状态实现3D点云识别的迭代决策,在ModelNet等数据集取得结果,机制可迁移至密集预测,能耗可降2.8倍至1.35倍。

Comments 28 pages, 9 figures, 17 tables. Supplementary material included as appendices A-J

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.11259 2026-08-25 math.ST 版本更新 50%

On the stability of filtration functions for dependent data with applications to break detection

关于相依数据的过滤函数稳定性及其在断点检测中的应用

Johannes Krebs, Daniel Rademacher

专题命中 点云 :point cloud(abstract)

AI总结 本文研究拓扑数据分析中过滤函数的稳定性,据此开发检测弱相依数据拓扑对象序列结构断点的检验方法,可应用于伯努利移位系统的持久性图统计。

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 3D生成 1 篇

2608.19567 2026-08-25 cs.CV 版本更新 79%

Block3D: Efficient Text-to-3D Generation via Block-Wise Diffusion

Block3D:基于分块扩散的高效文本到3D生成

Bowen Cui, Weijie Wang, Zeyu Zhang, Yefei He, Mingda Lin, Haoyu Zhao, Yuanyu He, Donny Y. Chen, Feng Chen, Bohan Zhuang

机构 * ZipLab, Zhejiang University(浙江大学ZipLab) University of California, Berkeley(加州大学伯克利分校) Wuhan University(武汉大学) Monash University(莫纳什大学) University of Adelaide(阿德莱德大学)

专题命中 3D生成 :3D generation(title,abstract);分类 cs.CV

AI总结 针对文本到3D生成成本高的问题,提出Block3D分块扩散框架,通过分块生成、联合去噪及置信度引导的块内修正,在保持几何保真度的同时实现5.15倍的速度提升。

Comments Project page: this https URL (https://alexandertsui.github.io/block3d/)

详情

展开后加载摘要…

URL PDF HTML 收藏

6. 空间理解 1 篇

2601.17885 2026-08-25 cs.CV cs.AI cs.RO 版本更新 62%

PEAfowl: Perception-Enhanced Multi-View Vision-Language-Action for Bimanual Manipulation

PEAfowl:感知增强的多视角视觉-语言-动作用于双臂操作

Qingyu Fan, Zhaoxiang Li, Jinrui Hu, Yi Lu, Wang Chen, Qiu Shen, Xiao-xiao Long, Yinghao Cai, Tao Lu, Shuo Wang, Xun Cao

机构 * Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) Nanjing University(南京大学)

专题命中 空间理解 :spatial understanding(abstract);分类 cs.CV、cs.RO

AI总结 PEAfowl通过增强感知的多视角视觉-语言-动作策略,提升双臂操作在复杂环境中的稳定性和成功率。

Comments Accepted by IEEE Robotics and Automation Letters (RA-L), 2026. This version includes an extended appendix with additional implementation details, deployment analysis, robustness evaluation, and real-world evaluation protocol. DOI: https://doi.org/10.1109/LRA.2026.3726379

详情

展开后加载摘要…

URL PDF HTML 收藏

7. SLAM与定位 2 篇

2604.15052 2026-08-25 cs.RO 版本更新 57%

CAVERS: Multimodal SLAM Data from a Natural Karstic Cave with Ground Truth Motion Capture

CAVERS:从自然喀斯特洞穴获取多模态SLAM数据并采用地面真实运动捕捉

Giacomo Franchini, David Rodríguez-Martínez, Alfonso Martínez-Petersen, C. J. Pérez-del-Pulgar, Marcello Chiaberge

机构 * Polytechnic of Turin Interdepartmental Centre for Service Robotics (PIC4SeR)(都灵理工大学服务机器人跨部门研究中心(PIC4SeR)) Systems Engineering and Automation Department, Universidad de Málaga(马德里大学系统工程与自动化系)

专题命中 SLAM与定位 :3D reconstruction(abstract);分类 cs.RO

AI总结 本文提出CAVERS数据集,通过地面真实运动捕捉系统提供高精度姿态和速度数据,验证了多模态传感器在复杂洞穴环境中的SLAM算法性能。

Comments 8 pages, 4 figures, accepted version

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.13183 2026-08-25 cs.CV cs.MM 版本更新 57%

GeoLink: A 3D-aware Framework to Improve Generalization for Cross-view Geo-localization

GeoLink:一种面向跨视角地理定位更好泛化的3D-aware框架

Hongyang Zhang, Yinhao Liu, Haitao Zhang, Zhongyi Wen, Zhenyu Kuang, Shuxian Liang, Xian-Sheng Hua

机构 * CUHK(SZ)(香港中文大学(深圳)) XMU(厦门大学) UESTC(电子科技大学) Foshan University(佛山大学) Tongji University(同济大学)

专题命中 SLAM与定位 :point cloud(abstract);分类 cs.CV

AI总结 本文提出GeoLink框架,通过3D感知的语义一致方法提升跨视角地理定位的泛化能力,实验表明其在多个基准测试中优于现有方法,能适应未见领域和多变天气。

详情

展开后加载摘要…

URL PDF HTML 收藏

8. 其他3D视觉 1 篇

2607.14935 2026-08-25 cs.CV 版本更新 70%

VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding

VideoChat3:用于高效通用视频理解的全开放视频多模态语言模型

Xinhao Li, Yuhan Zhu, Xiangyu Zeng, Yuhao Dong, Haoning Wu, Zhiqiu Zhang, Yuandong Yang, Changlian Ma, Qingyu Zhang, Yansong Shi, Xinyu Chen, Haoran Chen, Zizheng Huang, Jun Zhang, Kun Ouyang, Lin Sui, Ziang Yan, Yicheng Xu, Chenting Wang, Yinan He, Hongjie Zhang, Yi Wang, Yu Qiao, Yali Wang, Ziwei Liu, Kai Chen, Limin Wang

机构 * Nanjing University(南京大学) Shanghai AI Laboratory(上海人工智能实验室) Nanyang Technological University(南洋理工大学) Peking University(北京大学)

专题命中 其他3D视觉 :3D vision(abstract,abstract_cn);分类 cs.CV

AI总结 研究针对视频理解开源模型局限,提出VideoChat3。通过I3D-ViT等提升效率,利用可扩展视频数据合成管道生成训练数据集提升泛化性,以4B参数在多基准测试中超越同等或更多参数的开源模型,实现泛化与计算效率平衡。

详情

展开后加载摘要…

URL PDF HTML 收藏