arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

2026-02-17 至 2026-02-17 共收录 22 信号源:cs.CV, cs.GR, cs.RO

1. 三维重建 5 篇

2602.13909 2026-02-17 cs.RO cs.CV 84%

High-fidelity 3D reconstruction for planetary exploration

高保真行星探测的3D重建

Alfonso Martínez-Petersen, Levin Gerdes, David Rodríguez-Martínez, C. J. Pérez-del-Pulgar

专题命中 三维重建 :3D reconstruction(title);NeRF(abstract);Gaussian Splatting(abstract);分类 cs.CV、cs.RO

AI总结 本文提出了一种结合NeRF和高斯点绘的自动化3D重建管道,用于在极端环境下实现高保真的行星探测

Comments 7 pages, 3 figures, conference paper

Journal ref IEEE Conference on Artificial Intelligence (CAI) 2026, Special Session on AI for Space Exploration

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14119 2026-02-17 cs.CV 79%

GeoFusionLRM: Geometry-Aware Self-Correction for Consistent 3D Reconstruction

GeoFusionLRM: 基于几何的自我校正以实现一致的3D重建

Ahmet Burak Yildirim, Tuna Saygin, Duygu Ceylan, Aysegul Dundar

机构 * Bilkent University(比尔肯特大学) Adobe Research(Adobe研究)

专题命中 三维重建 :3D reconstruction(title,abstract);分类 cs.CV

AI总结 GeoFusionLRM通过反馈几何信息提升单图像3D重建的几何精度与一致性

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13801 2026-02-17 cs.CV 70%

Joint Orientation and Weight Optimization for Robust Watertight Surface Reconstruction via Dirichlet-Regularized Winding Fields

联合方向与权重优化用于通过狄利克雷正则化缠绕场实现稳健的紧致表面重建

Jiaze Li, Daisheng Jin, Fei Hou, Junhui Hou, Zheng Liu, Shiqing Xin, Wenping Wang, Ying He

机构 * Nanyang Technological University(南洋理工大学) Institute of Software, Chinese Academy of Sciences(中国科学院软件研究所) University of Chinese Academy of Sciences(中国科学院大学) City University of Hong Kong(香港城市大学) China University of Geosciences (Wuhan)(中国地质大学(武汉)) Shandong University(山东大学) Texas A&M University(德克萨斯大学)

专题命中 三维重建 :Gaussian Splatting(abstract);point cloud(abstract);分类 cs.CV

AI总结 DiWR通过联合优化点方向、权重和置信系数,实现从非均匀采样点云中稳健重建紧致表面。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14228 2026-02-17 cs.CV 57%

Learning Significant Persistent Homology Features for 3D Shape Understanding

学习显著持久同调特征以理解3D形状

Prachi Kudeshia, Jiju Poovvancheri

机构 * Saint Mary's University(圣玛丽大学)

专题命中 三维重建 :point cloud(abstract);分类 cs.CV

AI总结 本文提出TopoGAT方法,通过学习显著持久同调特征提升3D形状分析的性能。

Comments 17 pages, 10 figures, Preprint under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00220 2026-02-17 eess.IV cs.CV 57%

Deep learning Based Correction Algorithms for 3D Medical Reconstruction in Computed Tomography and Macroscopic Imaging

基于深度学习的3D医学重建在计算机断层扫描和宏观成像中的校正算法

Tomasz Les, Tomasz Markiewicz, Malgorzata Lorent, Miroslaw Dziekiewicz, Krzysztof Siwek

机构 * University of Technology(技术大学) Military Institute of Medicine(军事医学研究院) Institute of Tuberculosis and Lung Diseases(肺结核和肺病研究所)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

AI总结 本文提出了一种混合两阶段配准框架,结合几何先验和深度学习,提升3D医学重建的精度和解剖真实感。

Comments 23 pages, 9 figures, submitted to Applied Sciences (MDPI)

详情

展开后加载摘要…

URL PDF HTML 收藏

2. Gaussian Splatting 5 篇

2506.03407 2026-02-17 cs.GR cs.AI cs.CV cs.LG 84%

Multi-Spectral Gaussian Splatting with Neural Color Representation

多光谱高斯点云渲染与神经颜色表示

Lukas Meyer, Josef Grün, Maximilian Weiherer, Bernhard Egger, Marc Stamminger, Linus Franke

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract);分类 cs.CV、cs.GR

AI总结 本文提出MS-Splatting框架,通过神经颜色表示实现多光谱3D高斯点云渲染,提升多光谱和单光谱渲染质量,应用于植被指数渲染。

Comments for project page, see https://meyerls.github.io/ms_splatting

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13549 2026-02-17 cs.CV 83%

Nighttime Autonomous Driving Scene Reconstruction with Physically-Based Gaussian Splatting

夜间自动驾驶场景重建与基于物理的高斯点云

Tae-Kyeong Kim, Xingxin Chen, Guile Wu, Chengjie Huang, Dongfeng Bai, Bingbing Liu

机构 * Huawei Noah’s Ark Lab(华为诺亚实验室) University of Toronto(多伦多大学)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract);分类 cs.CV

AI总结 本文提出基于物理的高斯点云方法,提升自动驾驶夜间场景重建质量,实现实时渲染并优于现有方法。

Comments ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14493 2026-02-17 cs.CV cs.GR 79%

Gaussian Mesh Renderer for Lightweight Differentiable Rendering

轻量可微渲染的高斯网格渲染器

Xinpeng Liu, Fumio Okura

机构 * The University of Osaka(大阪大学)

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);3DGS(abstract);novel view synthesis(abstract);分类 cs.CV、cs.GR

AI总结 提出一种轻量可微网格渲染器,利用3DGS高效光栅化过程,实现更平滑的梯度优化,提升视图合成的效率与质量。

Comments IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP 2026). GitHub: https://github.com/huntorochi/Gaussian-Mesh-Renderer

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13444 2026-02-17 cs.RO cs.AI cs.CV 73%

FlowHOI: Flow-based Semantics-Grounded Generation of Hand-Object Interactions for Dexterous Robot Manipulation

FlowHOI: 基于流的语义引导的双手-物体交互生成用于灵巧机器人操作

Huajian Zeng, Lingyun Chen, Jiaqi Yang, Yuantai Zhang, Fan Shi, Peidong Liu, Xingxing Zuo

机构 * Mohamed bin Zayed University of Artificial Intelligence(莫扎德·扎耶德人工智能大学) Technical University of Munich(慕尼黑技术大学) National University of Singapore(新加坡国立大学) Westlake University(西湖大学)

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);3DGS(abstract);分类 cs.CV、cs.RO

AI总结 FlowHOI通过语义引导的流匹配框架生成双手-物体交互序列,提升机器人操作的准确性和效率。

Comments Project Page: https://huajian-zeng.github.io/projects/flowhoi/

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14929 2026-02-17 cs.CV 57%

Wrivinder: Towards Spatial Intelligence for Geo-locating Ground Images onto Satellite Imagery

Wrivinder:迈向空间智能:将地面图像定位到卫星图像

Chandrakanth Gudavalli, Tajuddin Manhar Mohammed, Abhay Yadav, Ananth Vishnu Bhaskar, Hardik Prajapati, Cheng Peng, Rama Chellappa, Shivkumar Chandrasekaran, B. S. Manjunath

机构 * Mayachitra, Inc.(Mayachitra公司) Johns Hopkins University(约翰霍普金斯大学)

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);分类 cs.CV

AI总结 Wrivinder通过聚合地面影像与卫星图像,实现零样本的几何驱动定位,提供首个全面的跨视角对齐基准。

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 点云 5 篇

2512.09407 2026-02-17 cs.CV 80%

Geometry-to-Image Synthesis-Driven Generative Point Cloud Registration

基于几何到图像合成的生成式点云配准

Haobo Jiang, Jin Xie, Jian Yang, Liang Yu, Jianmin Zheng

机构 * ANGEL CorpLab and College of Computing and Data Science, Nanyang Technological University, Singapore(ANGEL CorpLab和计算与数据科学学院,南洋理工大学,新加坡) Alibaba Cloud, Alibaba Group, China(阿里巴巴云,阿里巴巴集团,中国) PCA Lab, VCIP, College of Computer Science, Nankai University, China(PCA实验室,VCIP,计算机科学学院,南开大学,中国) State Key Laboratory for Novel Software Technology & Schoolof Intelligence Science and Technology, Nanjing University, China(新型软件技术国家重点实验室与智能科学与技术学校,南京大学,中国)

专题命中 点云 :point cloud(title,abstract);分类 cs.CV

AI总结 本文提出生成式点云配准方法,通过生成跨视角一致的图像对,结合几何与颜色特征融合,提升3D配准性能。

Comments Journal extension of the ICML 2025 paper "Generative Point Cloud Registration". This version adopts a new title, and includes substantial methodological improvements, additional experiments, and extended analysis. Under review at IEEE TPAMI

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14193 2026-02-17 cs.RO cs.CV cs.LG 62%

Learning Part-Aware Dense 3D Feature Field for Generalizable Articulated Object Manipulation

学习具有部分感知的密集3D特征场以实现通用的可变形物体操作

Yue Chen, Muqing Jiang, Kaifeng Zheng, Jiaqi Liang, Chenrui Tie, Haoran Lu, Ruihai Wu, Hao Dong

机构 * Peking University(北京大学) Beijing Institute of Technology(北京理工大学) National University of Singapore(新加坡国立大学)

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

AI总结 本文提出PA3FF,一种具有部分感知的密集3D特征场,用于提升可变形物体操作的泛化能力,通过对比学习训练,实现高效且通用的机器人操作

Comments Accept to ICLR 2026, Project page: https://pa3ff.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14525 2026-02-17 cs.CV 57%

Cross-view Domain Generalization via Geometric Consistency for LiDAR Semantic Segmentation

通过几何一致性实现跨视角领域泛化的LiDAR语义分割

Jindong Zhao, Yuan Gao, Yang Xia, Sheng Nie, Jun Yue, Weiwei Sun, Shaobo Xia

机构 * School of Aeronautic Engineering, Changsha University of Science and Technology(长沙理工大学航空工程学院) Aerospace Information Research Institute, Chinese Academy of Sciences(中国科学院航空信息研究所) School of Automation, Central South University(中南大学自动化学院) University of British Columbia(不列颠哥伦比亚大学)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 CVGC通过几何一致性模块解决跨视角LiDAR语义分割中的领域泛化问题,优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13844 2026-02-17 cs.CV 57%

Synthetic Dataset Generation and Validation for Robotic Surgery Instrument Segmentation

用于机器人手术器械分割的合成数据集生成与验证

Giorgio Chiesa, Rossella Borra, Vittorio Lauro, Sabrina De Cillis, Daniele Amparore, Cristian Fiori, Riccardo Renzulli, Marco Grangetto

机构 * University of Turin, Italy(都灵大学)

专题命中 点云 :3D reconstruction(abstract);分类 cs.CV

AI总结 本文提出了一种生成和验证用于机器人手术器械分割的合成数据集的方法,通过自动化流程生成逼真数据,并验证其在模型训练中的有效性。

Comments Accepted at ISBI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.13010 2026-02-17 cs.LG cs.CE physics.comp-ph stat.ML 50%

A Resolution Independent Neural Operator

独立于分辨率的神经算子

Bahador Bahmani, Somdatta Goswami, Ioannis G. Kevrekidis, Michael D. Shields

机构 * Hopkins Extreme Materials Institute(霍普金斯极端材料研究所) Dept. of Civil and Systems Engineering(土木与系统工程系) Johns Hopkins University(约翰·霍普金斯大学) Dept. of Chemical and Biomolecular Engineering(化学与生物分子工程系) Dept. of Applied Mathematics and Statistics(应用数学与统计学系)

专题命中 点云 :point cloud(abstract)

AI总结 本文提出了一种独立于分辨率的神经算子架构RINO,通过自适应学习连续基函数,实现对任意采样输入输出函数的高效处理。

Journal ref Journal of Computational Physics, 539, 2025, 114233

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 3D生成 3 篇

2602.13818 2026-02-17 cs.CV cs.LG 79%

VAR-3D: View-aware Auto-Regressive Model for Text-to-3D Generation via a 3D Tokenizer

VAR-3D: 基于视图感知的自回归模型用于文本到3D生成 via 3D标记器

Zongcheng Han, Dongyan Cao, Haoran Sun, Yu Hong

机构 * School of Computer Science and Technology(计算机科学与技术学院) Soochow University(苏州大学) Harbin Institute of Technology(哈尔滨工业大学)

专题命中 3D生成 :3D generation(title,abstract);分类 cs.CV

AI总结 VAR-3D通过视图感知的3D标记器和渲染监督训练策略,提升文本到3D生成的质量和对齐性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.04641 2026-02-17 cs.CV cs.AI cs.LG 57%

Simulating the Real World: A Unified Survey of Multimodal Generative Models

模拟现实世界:多模态生成模型的统一综述

Yuqi Hu, Longguang Wang, Xian Liu, Ling-Hao Chen, Yuwei Guo, Yukai Shi, Ce Liu, Anyi Rao, Zeyu Wang, Hui Xiong

机构 * Thrust of Artificial Intelligence, The Hong Kong University of Science and Technology (Guangzhou)(人工智能前沿技术研究所,香港科学与技术大学(广州)) Department of Computer Science and Engineering, The Hong Kong University of Science and Technology Hong Kong SAR(计算机科学与工程系,香港科学与技术大学香港特别行政区) MMLab, The Hong Kong University of Science and Technology(多模态实验室,香港科学与技术大学) School of Electronics and Communication Engineering, Shenzhen Campus of Sun Yat-sen University(电子与通信工程学院,中山大学深圳校区) The Chinese University of Hong Kong, Hong Kong, China(香港中文大学,香港,中国) Tsinghua University, Guangdong, China(清华大学,广东,中国) Bosch (China) Investment Co., Ltd., Shanghai, China(博世(中国)投资有限公司,上海,中国)

专题命中 3D生成 :3D generation(abstract);分类 cs.CV

AI总结 本文首次系统性地统一研究了2D、视频、3D和4D生成,为多模态生成模型和现实世界模拟提供了统一框架的综述。

Comments Repository for the related papers at https://github.com/ALEEEHU/World-Simulator

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.09713 2026-02-17 cs.CV 57%

Stroke3D: Lifting 2D strokes into rigged 3D model via latent diffusion models

Stroke3D: 通过潜在扩散模型将2D笔触提升为拟合的3D模型

Ruisi Zhao, Haoren Zheng, Zongxin Yang, Hehe Fan, Yi Yang

机构 * ReLER, CCAI, Zhejiang University(ReLER、CCAAI、浙江大学) DBMI, HMS, Harvard University(DBMI、HMS、哈佛大学)

专题命中 3D生成 :3D generation(abstract);分类 cs.CV

AI总结 Stroke3D通过潜在扩散模型,将用户绘制的2D笔触和文本提示转化为拟合的3D模型,实现可控的骨骼生成和纹理网格合成。

Comments Accepted by ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 空间理解 2 篇

2602.13588 2026-02-17 cs.CV cs.AI 57%

Two-Stream Interactive Joint Learning of Scene Parsing and Geometric Vision Tasks

双流交互式场景解析与几何视觉任务联合学习

Guanfeng Tang, Hongbo Zhao, Ziwei Long, Jiayao Li, Bohong Xiao, Wei Ye, Hanli Wang, Rui Fan

机构 * College of Electronic and Information Engineering, Tongji University(同济大学电子信息学院) Shanghai Research Institute for Intelligent Autonomous Systems, Tongji University(同济大学智能自主系统研究所) College of Computer Science and Technology, Tongji University(同济大学计算机科学与技术学院) Department of Vehicle Control System and Software Development, NIO(蔚来汽车车辆控制系统与软件开发部门) Department of Automotive Engineering, Jilin University(吉林大学汽车工程学院) Key Laboratory of Embedded System and Service Computing (Ministry of Education), Tongji University(同济大学嵌入式系统与服务计算重点实验室) College of Electronic and Information Engineering, Shanghai Institute of Intelligent Science and Technology(上海智能科学与技术研究院电子信息学院) Shanghai Key Laboratory of Intelligent Autonomous Systems(上海智能自主系统重点实验室)

专题命中 空间理解 :spatial understanding(abstract);分类 cs.CV

AI总结 TwInS通过双流交互式联合学习框架,实现场景解析与几何视觉任务的同步优化,采用跨任务适配器提升性能并减少对人工标注的依赖。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14442 2026-02-17 cs.HC 50%

Touching Movement: 3D Tactile Poses for Supporting Blind People in Learning Body Movements

触觉运动:3D触觉姿势用于支持视障人士学习身体运动

Kengo Tanaka, Xiyue Wang, Hironobu Takagi, Yoichi Ochiai, Chieko Asakawa

专题命中 空间理解 :spatial understanding(abstract)

AI总结 本研究通过3D打印人体模型帮助视障人士更有效地学习身体运动,实验显示该方法在理解速度、准确性及学习动机方面优于传统教学方法。

Comments Accepted to TEI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏

6. SLAM与定位 2 篇

2602.14662 2026-02-17 cs.CV cs.RO 81%

Advances in Global Solvers for 3D Vision

3D视觉中全局求解器的进展

Zhenjun Zhao, Heng Yang, Bangyan Liao, Yingping Zeng, Shaocheng Yan, Yingdong Gu, Peidong Liu, Yi Zhou, Haoang Li, Javier Civera

专题命中 SLAM与定位 :3D vision(title,abstract);分类 cs.CV、cs.RO

AI总结 本文系统回顾了3D视觉中全局求解器的发展,探讨了三种核心范式及其在非凸优化问题中的应用与挑战。

Comments Comprehensive survey; 37 pages, 7 figures, 3 tables. Project page with literature tracking and code tutorials: https://github.com/ericzzj1989/Awesome-Global-Solvers-for-3D-Vision

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15990 2026-02-17 cs.RO cs.CV 62%

GelSLAM: A Real-time, High-Fidelity, and Robust 3D Tactile SLAM System

GelSLAM:一种实时、高保真度和鲁棒的3D触觉SLAM系统

Hung-Jui Huang, Mohammad Amin Mirzaee, Michael Kaess, Wenzhen Yuan

机构 * Carnegie Mellon University(卡内基梅隆大学) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 SLAM与定位 :point cloud(abstract);分类 cs.CV、cs.RO

AI总结 GelSLAM通过触觉传感实现实时高保真3D SLAM,以高精度重建物体形状并提升手部操作任务的鲁棒性。

Comments 20 pages

详情

展开后加载摘要…

URL PDF HTML 收藏