arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

2025-12-03 至 2025-12-03 共收录 22 信号源:cs.CV, cs.GR, cs.RO

1. 三维重建 5 篇

2503.14537 2025-12-03 cs.CV cs.RO 81%

Learning-based 3D Reconstruction in Autonomous Driving: A Comprehensive Survey

基于学习的自动驾驶3D重建:全面综述

Liewen Liao, Weihao Yan, Wang Xu, Ming Yang, Songan Zhang, H. Eric Tseng

机构 * Global Institute of Future Technology, Shanghai Jiao Tong University(未来技术全球研究院,上海交通大学) School of Automation and Intelligent Sensing, Shanghai Jiao Tong University(自动化与智能感知学院,上海交通大学) Department of Electrical Engineering, University of Texas at Arlington(电气工程系,德克萨斯理工大学)

专题命中 三维重建 :3D reconstruction(title,abstract);分类 cs.CV、cs.RO

AI总结 本文综述了基于学习的3D重建在自动驾驶中的应用,分析了技术发展和挑战,提出了未来研究方向。

Comments Published in IEEE Trans. on Intelligent Transportation Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22052 2025-12-03 cs.CV 79%

Ov3R: Open-Vocabulary Semantic 3D Reconstruction from RGB Videos

Ov3R:从RGB视频流中进行开放词汇语义3D重建

Ziren Gong, Xiaohan Li, Fabio Tosi, Jiawei Han, Stefano Mattoccia, Jianfei Cai, Matteo Poggi

机构 * University of Bologna(博洛尼亚大学) USTC(中科大) BIT(北京理工) Monash University(墨尔本大学)

专题命中 三维重建 :3D reconstruction(title,abstract);分类 cs.CV

AI总结 Ov3R通过结合CLIP语义和融合描述符,实现从RGB视频流中进行开放词汇语义3D重建,提升空间人工智能的实时性和语义感知能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02375 2025-12-03 cs.CV 70%

On-the-fly Feedback SfM: Online Explore-and-Exploit UAV Photogrammetry with Incremental Mesh Quality-Aware Indicator and Predictive Path Planning

实时反馈结构光束法重建:面向无人机摄影测量的在线探索与利用方法,结合增量网格质量感知指标与预测路径规划

Liyuan Lou, Wanyun Li, Wentian Gan, Yifei Yu, Tengfei Wang, Xin Wang, Zongqian Zhan

机构 * School of Geodesy and Geomatics, Wuhan University(测绘学院,武汉大学)

专题命中 三维重建 :3D reconstruction(abstract);point cloud(abstract);分类 cs.CV

AI总结 本文提出实时反馈结构光束法重建方法,通过增量网格质量感知和预测路径规划,实现无人机摄影测量的在线探索与利用,提升实时重建与反馈效率。

Comments This work was submitted to IEEE GRSM Journal for consideration.COPYRIGHT would be transferred once it get accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.22686 2025-12-03 cs.CV 70%

Emergent Extreme-View Geometry in 3D Foundation Models

三维基础模型中的涌现极端视角几何

Yiwen Zhang, Joseph Tung, Ruojin Cai, David Fouhey, Hadar Averbuch-Elor

机构 * Cornell University(康奈尔大学) New York University(纽约大学) Kempner Institute, Harvard University(哈佛大学凯普勒研究所)

专题命中 三维重建 :3D vision(abstract);3D reconstruction(abstract);分类 cs.CV

AI总结 本文研究了三维基础模型在极端视角下的几何理解能力,并提出轻量级对齐方案提升其相对姿态估计性能,同时引入新的未见过的互联网场景基准。

Comments Project page is at https://ext-3dfms.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02263 2025-12-03 cs.HC cs.GR 57%

DepthScape: Authoring 2.5D Designs via Depth Estimation, Semantic Understanding, and Geometry Extraction

DepthScape: 通过深度估计、语义理解和几何提取进行2.5D设计创作

Xia Su, Cuong Nguyen, Matheus A. Gadelha, Jon E. Froehlich

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.GR

AI总结 DepthScape通过深度估计、语义理解和几何提取,实现2.5D设计创作,提升视觉真实感与动态效果。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. Gaussian Splatting 7 篇

2512.02932 2025-12-03 cs.CV cs.AI 89%

EGGS: Exchangeable 2D/3D Gaussian Splatting for Geometry-Appearance Balanced Novel View Synthesis

EGGS: 可交换的2D/3D高斯点扩散用于几何-外观平衡的新视角合成

Yancheng Zhang, Guangyu Sun, Chen Chen

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);novel view synthesis(title,abstract);3DGS(abstract);分类 cs.CV

AI总结 EGGS通过融合2D和3D高斯点扩散,平衡外观与几何精度,提升新视角合成的渲染质量与效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02664 2025-12-03 cs.CV 87%

PolarGuide-GSDR: 3D Gaussian Splatting Driven by Polarization Priors and Deferred Reflection for Real-World Reflective Scenes

PolarGuide-GSDR: 由偏振先验和延迟反射驱动的3D高斯散射

Derui Shan, Qian Qiao, Hao Lu, Tao Du, Peng Lu

机构 * North China University of Technology(中国北方工业大学) Beijing University of Posts and Telecommunications(北京邮电大学) University of Science and Technology Beijing(北京科技大学)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);NeRF(abstract);3DGS(abstract);novel view synthesis(abstract)

AI总结 PolarGuide-GSDR通过偏振引导机制改进3DGS,实现高保真度反射分离与全场景重建,无需环境映射或材料假设,提升实时性能与可解释性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.03059 2025-12-03 cs.CV cs.LG 85%

Compressing 3D Gaussian Splatting by Noise-Substituted Vector Quantization

通过噪声替代向量量化压缩3D高斯散射

Haishan Wang, Mohammad Hassan Vali, Arno Solin

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract);3D reconstruction(abstract);分类 cs.CV

AI总结 本文提出通过噪声替代向量量化技术压缩3D高斯散射,有效减少内存消耗并保持重建质量,适用于实际应用。

Comments Appearing in Scandinavian Conference on Image Analysis (SCIA) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02621 2025-12-03 cs.CV cs.GR 84%

Content-Aware Texturing for Gaussian Splatting

具有内容感知的高斯点散布纹理生成

Panagiotis Papantonakis, Georgios Kopanas, Fredo Durand, George Drettakis

机构 * Inria(法国国家信息与自动化技术研究院) Université Côte D’Azur(蔚蓝海岸大学) Google(谷歌) Runway ML(Runway机器学习) MIT CSAIL(麻省理工学院计算机科学与人工智能实验室)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3D reconstruction(abstract);分类 cs.CV、cs.GR

AI总结 本文提出了一种基于内容感知的高斯点散布纹理生成方法,通过自适应调整纹理分辨率来提高图像质量和参数效率。

Comments Project Page: https://repo-sam.inria.fr/nerphys/gs-texturing/

Journal ref Eurographics Symposium on Rendering (Symposium Track), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.01846 2025-12-03 cs.CV 83%

UVGS: Reimagining Unstructured 3D Gaussian Splatting using UV Mapping

UVGS: 重新构想使用UV映射的无结构3D高斯点撒技术

Aashish Rai, Dilin Wang, Mihir Jain, Nikolaos Sarafianos, Kefan Chen, Srinath Sridhar, Aayush Prakash

机构 * Brown University(布朗大学) Meta Reality Labs(Meta现实实验室)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract);分类 cs.CV

AI总结 本文提出UVGS,通过UV映射将无结构3D高斯点撒转化为结构化2D表示,利用现有2D模型高效建模3D数据,并实现可扩展的生成应用。

Comments https://ivl.cs.brown.edu/uvgs

Journal ref CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02443 2025-12-03 cs.GR cs.CV 82%

PRIMU: Uncertainty Estimation for Novel Views in Gaussian Splatting from Primitive-Based Representations of Error and Coverage

PRIMU:基于原始体表示的误差与覆盖度的不确定性估计

Thomas Gottwald, Edgar Heinert, Peter Stehr, Chamuditha Jayanga Galappaththige, Matthias Rottmann

机构 * Department of Mathematics(数学系) University of Wuppertal(乌珀塔尔大学) Institute of Computer Science(计算机科学研究所) University of Osnabrück(奥斯纳布吕克大学) Centre for Robotics(机器人中心) Queensland University of Technology(昆士兰理工大学)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);分类 cs.CV、cs.GR

AI总结 PRIMU通过基于原始体的误差和覆盖度表示,实现了高斯点扩散中对新视图不确定性的有效估计,优于现有方法,尤其在深度和前景物体估计上表现突出。

Comments Revised writing and figures; additional Gaussian Splatting experiments; added baselines and datasets; active view-selection experiments

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08294 2025-12-03 cs.CV 57%

SkelSplat: Robust Multi-view 3D Human Pose Estimation with Differentiable Gaussian Rendering

SkelSplat: 基于可微高斯渲染的鲁棒多视角3D人体姿态估计

Laura Bragagnolo, Leonardo Barcellona, Stefano Ghidoni

机构 * University of Padova(帕多瓦大学) University of Amsterdam(阿姆斯特丹大学)

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);分类 cs.CV

AI总结 SkelSplat通过可微高斯渲染实现多视角3D人体姿态估计,无需3D真实数据监督,有效提升跨数据集泛化能力与遮挡鲁棒性。

Comments WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 点云 5 篇

2512.02972 2025-12-03 cs.CV cs.RO 62%

BEVDilation: LiDAR-Centric Multi-Modal Fusion for 3D Object Detection

BEVDilation:以LiDAR为中心的多模态融合用于3D目标检测

Guowen Zhang, Chenhang He, Liyi Chen, Lei Zhang

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

AI总结 BEVDilation提出以LiDAR为中心的多模态融合方法,通过稀疏体素扩张和语义引导BEV扩张模块提升3D目标检测性能,有效缓解深度误差带来的空间错位问题。

Comments Accept by AAAI26

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02394 2025-12-03 cs.CV 57%

Reproducing and Extending RaDelft 4D Radar with Camera-Assisted Labels

重现并扩展RaDelft 4D雷达与摄像头辅助标签

Kejia Hu, Mohammed Alsakabi, John M. Dolan, Ozan K. Tonguz

机构 * Department of Electrical and Computer Engineering, College of Engineering(电气与计算机工程系) The Robotics Institute, School of Computer Science(机器人研究所)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本研究重现并扩展RaDelft 4D雷达数据集,通过摄像头引导的标签生成方法实现雷达点云的高精度标注,并分析雾度对雷达标签性能的影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05858 2025-12-03 cs.RO 57%

ViTaMIn-B: A Reliable and Efficient Visuo-Tactile Bimanual Manipulation Interface

ViTaMIn-B: 一种可靠且高效的视觉-触觉双臂操作接口

Chuanyu Li, Chaoyi Liu, Daotan Wang, Shuyu Zhang, Lusong Li, Zecui Zeng, Fangchen Liu, Jing Xu, Rui Chen

机构 * Tsinghua University(清华大学) University of California, Berkeley(加州大学伯克利分校) JD Explore Academy(JD探索学院) The Hong Kong Polytechnic University(香港理工大学)

专题命中 点云 :point cloud(abstract);分类 cs.RO

AI总结 ViTaMIn-B通过DuoTact传感器和6自由度双臂位姿采集技术,实现了高效可靠的双臂操作任务数据采集。

Comments Project page: https://chuanyune.github.io/ViTaMIn-B_page/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07561 2025-12-03 cs.CV 57%

Alligat0R: Pre-Training Through Co-Visibility Segmentation for Relative Camera Pose Regression

Alligat0R:通过共视分割进行预训练以实现相对相机姿态回归

Thibaut Loiseau, Guillaume Bourmaud, Vincent Lepetit

机构 * LIGM, Ecole des Ponts, Univ. Gustave Eiffel, CNRS, France(LIGM,巴黎理工大学,埃菲尔大学,CNRS,法国) Univ. Bordeaux, CNRS, Bordeaux INP, IMS, UMR 5218, France(波尔多大学,CNRS,波尔多INP,IMS,UMR 5218,法国)

专题命中 点云 :3D reconstruction(abstract);分类 cs.CV

AI总结 Alligat0R通过共视分割预训练方法在相对相机姿态回归中优于CroCo

Comments NeurIPS 2025 Spotlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10341 2025-12-03 cs.LG 50%

Multi-View Graph Learning with Graph-Tuple

多视图图学习与图元组

Shiyu Chen, Ningyuan Huang, Soledad Villar

机构 * Department of Applied Mathematics and Statistics(应用数学与统计学系) Johns Hopkins University(约翰霍普金斯大学) Flatiron Institute(Flatiron研究所)

专题命中 点云 :point cloud(abstract)

AI总结 本文提出多视图图元组框架,通过异质信息传递架构提升图神经网络的表现力,应用于分子属性预测和宇宙学参数推断,展现其在处理密集图数据中的优势。

Comments Accepted as a poster at the TAG-DS 2025 Workshop (Topology, Algebra, and Geometry in Data Science). OpenReview: https://openreview.net/forum?id=s4ezAuj5xM

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 新视角合成 1 篇

2512.03045 2025-12-03 cs.CV 57%

CAMEO: Correspondence-Attention Alignment for Multi-View Diffusion Models

CAMEO:多视图扩散模型中的对应-注意力对齐

Minkyung Kwon, Jinhyeok Choi, Jiho Park, Seonghu Jeon, Jinhyuk Jang, Junyoung Seo, Minseop Kwak, Jin-Hwa Kim, Seungryong Kim

机构 * KAIST AI(韩国科学技术院人工智能研究中心) NAVER AI Lab(NAVER人工智能实验室) SNU AIIS(延世大学人工智能研究所)

专题命中 新视角合成 :novel view synthesis(abstract);分类 cs.CV

AI总结 CAMEO通过几何对应监督注意力图,提升多视图扩散模型的训练效率和生成质量。

Comments Project page: https://cvlab-kaist.github.io/CAMEO/

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 空间理解 2 篇

2503.21214 2025-12-03 cs.CV cs.CL 74%

VoxRep: Enhancing 3D Spatial Understanding in 2D Vision-Language Models via Voxel Representation

VoxRep:通过体素表示增强2D视觉-语言模型的3D空间理解

Alan Dao, Norapat Buppodom

机构 * Menlo Research(Menlo研究)

专题命中 空间理解 :spatial understanding(title);分类 cs.CV

AI总结 本文提出VoxRep方法,通过将体素空间切分为2D切片并输入预训练的视觉-语言模型,实现对3D环境的高效语义理解。

Journal ref Proc. APSIPA ASC 2025, pp. 1464-1469

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02715 2025-12-03 cs.CV 57%

GeoViS: Geospatially Rewarded Visual Search for Remote Sensing Visual Grounding

GeoViS: 基于地理奖励的视觉搜索用于遥感视觉定位

Peirong Zhang, Yidan Zhang, Luxiao Xu, Jinliang Lin, Zonghao Guo, Fengxiang Wang, Xue Yang, Kaiwen Wei, Lei Wang

机构 * Aerospace Information Research Institute, Chinese Academy of Sciences(中国科学院航空信息研究所) University of Chinese Academy of Sciences(中国科学院大学) Tsinghua University(清华大学) National University of Defense Technology(国防科技大学) Shanghai Jiao Tong University(上海交通大学) Chongqing University(重庆大学)

专题命中 空间理解 :spatial understanding(abstract);分类 cs.CV

AI总结 GeoViS通过地理奖励机制,实现了遥感影像中的细粒度视觉定位,通过逐步搜索和推理提升小目标检测精度与跨领域泛化能力。

Comments 11 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

6. SLAM与定位 1 篇

2512.02162 2025-12-03 cs.CV q-bio.QM 57%

Mapping of Lesion Images to Somatic Mutations

将病变图像映射到体细胞突变

Rahul Mehta

机构 * University of Illinois at Chicago(伊利诺伊大学香槟分校) University of Texas MD Anderson Cancer Center(德克萨斯大学MD安德森癌症中心)

专题命中 SLAM与定位 :point cloud(abstract);分类 cs.CV

AI总结 本文提出LLOST模型,通过点云表示和共享潜在空间,将医学影像与体细胞突变数据关联,以预测突变特征和癌症类型。

Comments https://dl.acm.org/doi/abs/10.1145/3340531.3414074#sec-terms

详情

展开后加载摘要…

URL PDF HTML 收藏

7. 其他3D视觉 1 篇

2511.11722 2025-12-03 cs.LG cs.AI cs.CV cs.SY eess.SY 57%

Fast 3D Surrogate Modeling for Data Center Thermal Management

快速三维代理建模用于数据中心热管理

Soumyendu Sarkar, Antonio Guillen-Perez, Zachariah J Carmichael, Avisek Naug, Refik Mert Cam, Vineet Gundecha, Ashwin Ramesh Babu, Sahand Ghorbanpour, Ricardo Luna Gutierrez

专题命中 其他3D视觉 :3D vision(abstract);分类 cs.CV

AI总结 本文提出基于视觉的快速三维代理建模方法,用于实时预测数据中心温度分布,实现高效的冷却控制与能耗降低。

Comments Submitted to AAAI 2026 Conference

详情

展开后加载摘要…

URL PDF HTML 收藏