arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

2026-01-08 至 2026-01-08 共收录 10 信号源:cs.CV, cs.GR, cs.RO

1. 三维重建 2 篇

2502.01045 2026-01-08 cs.CV cs.GR 62%

WonderHuman: Hallucinating Unseen Parts in Dynamic 3D Human Reconstruction

WonderHuman: 在动态3D人体重建中生成未见部分

Zilong Wang, Zhiyang Dou, Yuan Liu, Cheng Lin, Xiao Dong, Yunhui Guo, Chenxu Zhang, Xin Li, Wenping Wang, Xiaohu Guo

机构 * Department of Computer Science, The University of Texas at Dallas(德克萨斯大学达拉斯分校计算机科学系) Computer Graphics Group, The University of Hong Kong(香港大学计算机图形组) School of Engineering, The Hong Kong University of Science and Technology(香港科学与技术大学工程学院) Guangdong Provincial/Zhuhai Key Laboratory of IRADS, Beijing Normal-Hong Kong Baptist University(广东/珠海IRADS重点实验室,北京师范大学-香港 Baptist大学) Department of Computer Science & Engineering, Texas A&M University(德克萨斯A&M大学计算机科学与工程系)

专题命中 三维重建 :novel view synthesis(abstract);分类 cs.CV、cs.GR

AI总结 WonderHuman通过双空间优化和分数蒸馏采样技术,从单目视频实现高保真动态人体重建,尤其擅长生成未见人体部分。

Journal ref IEEE Transactions on Visualization and Computer Graphics, vol. 31, no. 12, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.15735 2026-01-08 math.NA cs.NA 50%

A Higher Order Local Mesh Method for Approximating 1-Laplacians on Unknown Manifolds

一种高阶局部网格方法用于在未知流形上近似1-拉普拉斯算子

John Wilson Peoples, John Harlim

专题命中 三维重建 :point cloud(abstract)

AI总结 本文提出一种高阶局部网格方法,用于在未知流形上近似1-拉普拉斯算子,通过结合局部网格和曲率信息实现高精度估计。

Comments 27 pages, 5 figure

详情

展开后加载摘要…

URL PDF HTML 收藏

2. Gaussian Splatting 1 篇

2601.03319 2026-01-08 cs.GR cs.AI cs.LG 83%

CaricatureGS: Exaggerating 3D Gaussian Splatting Faces With Gaussian Curvature

CaricatureGS: 通过高斯曲率夸大3D高斯点云人脸

Eldad Matmon, Amit Bracha, Noam Rotstein, Ron Kimmel

机构 * Technion – Israel Institute of Technology(技术学院–以色列理工学院)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract);分类 cs.GR

AI总结 CaricatureGS通过高斯曲率夸大技术,结合3D高斯点云和局部仿射变换,实现可控且逼真的卡通化虚拟形象生成。

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 点云 2 篇

2410.18987 2026-01-08 cs.CV cs.LG 79%

Point Cloud Synthesis Using Inner Product Transforms

利用内积变换进行点云合成

Ernst Röell, Bastian Rieck

机构 * AIDOS Lab, University of Fribourg(AIDOS实验室,弗里堡大学) Institute of AI for Health, Helmholtz Munich(健康人工智能研究所,海德堡慕尼黑) Technical University of Munich(慕尼黑技术大学)

专题命中 点云 :point cloud(title,abstract);分类 cs.CV

AI总结 本文提出了一种利用内积变换编码点云几何-拓扑特性,实现高效且高质量点云合成的方法。

Comments Accepted at the 39th Conference on Neural Information Processing Systems (NeurIPS) 2025. Our code is available at https://github.com/aidos-lab/inner-product-transforms

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.03308 2026-01-08 q-bio.QM 50%

A Comprehensive Database of Leaf Temperature, Water, and CO 2 Fluxes in Young Oil Palm Plants Across Diverse Climate Scenarios

全面的年轻油棕植物叶温度、水分和CO2通量数据库,在多样的气候情景下

Raphael Perez, Valentin Torrelli, Sandrine Roques, Sébastien Devidal, Clément Piel, Damien Landais, Merlin Ramel, Thomas Arsouze, Julien Lamour, Jean-Pierre Caliman, Rémi Vezy

专题命中 点云 :point cloud(abstract)

AI总结 本研究提供了一个全面的油棕植物叶温度、水分和CO2通量数据库,用于评估FSPM模型的准确性,促进生物物理模型的基准测试和改进。

Comments Data and code availability: The raw data and scripts used to generate the final database are detailed and accessible on Zenodo (Vezy et al., 2025), the code is also accessible via a Github repository (https://github.com/PalmStudio/Biophysics\_database\_palm), and we also provide a companion website (https://palmstudio.github.io/Biophysics\_database\_palm) showing how computations were made and the main results. The code to trigger the FLIR camera and for logging the precision scale data is also available on dedicated Zenodo repositories (Vezy, 2025a, 2025b)

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 新视角合成 1 篇

2601.03362 2026-01-08 cs.CV 70%

Guardians of the Hair: Rescuing Soft Boundaries in Depth, Stereo, and Novel Views

毛发守护者:恢复深度、立体和新视角中的柔软边界

Xiang Zhang, Yang Zhang, Lukas Mehl, Markus Gross, Christopher Schroers

专题命中 新视角合成 :3D vision(abstract);novel view synthesis(abstract);分类 cs.CV

AI总结 HairGuard通过深度修复网络和生成场景画家恢复3D视觉中柔软边界的细节,提升单目深度估计、立体转换和新视角合成的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 空间理解 2 篇

2601.03590 2026-01-08 cs.CV cs.AI 57%

Can LLMs See Without Pixels? Benchmarking Spatial Intelligence from Textual Descriptions

大语言模型能否通过像素感知空间智能?基于文本描述的空间智能基准测试

Zhongbin Guo, Zhen Yang, Yushan Li, Xinyue Zhang, Wenyu Gao, Jiacheng Wang, Chengzhi Li, Xiangrui Liu, Ping Jian

机构 * School of Computer Science & Technology, Beijing Institute of Technology(计算机科学与技术学院,北京理工大学)

专题命中 空间理解 :spatial understanding(abstract);分类 cs.CV

AI总结 SiT-Bench通过文本描述评估大语言模型的空间智能,揭示其在全局一致性方面的不足,并表明显式空间推理可提升性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.13683 2026-01-08 cs.CV 57%

I-Scene: 3D Instance Models are Implicit Generalizable Spatial Learners

I-Scene: 3D实例模型是隐式可泛化的空间学习者

Lu Ling, Yunhao Ge, Yichen Sheng, Aniket Bera

机构 * Purdue University(普渡大学) NVIDIA Research(英伟达研究)

专题命中 空间理解 :spatial understanding(abstract);分类 cs.CV

AI总结 I-Scene提出了一种基于3D实例模型的隐式空间学习方法,通过重新编程生成器实现跨场景的泛化能力,展示了其在空间推理和生成中的潜力。

详情

展开后加载摘要…

URL PDF HTML 收藏

6. SLAM与定位 1 篇

2601.03579 2026-01-08 cs.CV 57%

SpatiaLoc: Leveraging Multi-Level Spatial Enhanced Descriptors for Cross-Modal Localization

SpatiaLoc: 借助多级空间增强描述符进行跨模态定位

Tianyi Shang, Pengjie Xu, Zhaojun Deng, Zhenyu Li, Zhicong Chen, Lijun Wu

机构 * Fuzhou University(福州大学) Shandong Academy of Sciences(山东省科学院) Qingdao University(青岛大学) Tongji University(同济大学)

专题命中 SLAM与定位 :point cloud(abstract);分类 cs.CV

AI总结 SpatiaLoc通过多级空间增强描述符提升跨模态定位性能,采用粗到细策略结合贝塞尔曲线和频率域建模,实现更精确的机器人定位。

详情

展开后加载摘要…

URL PDF HTML 收藏

7. 其他3D视觉 1 篇

2601.03782 2026-01-08 cs.RO cs.AI cs.CV 62%

PointWorld: Scaling 3D World Models for In-The-Wild Robotic Manipulation

PointWorld: 为真实世界机器人操作扩展3D世界模型

Wenlong Huang, Yu-Wei Chao, Arsalan Mousavian, Ming-Yu Liu, Dieter Fox, Kaichun Mo, Li Fei-Fei

机构 * Stanford University(斯坦福大学) NVIDIA(英伟达)

专题命中 其他3D视觉 :3D vision(abstract);分类 cs.CV、cs.RO

AI总结 PointWorld通过统一状态和动作的3D点流模型,实现了在真实世界中机器人操作的高效预测与控制,无需额外演示或训练。

详情

展开后加载摘要…

URL PDF HTML 收藏